Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

1,189 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
PRAGMA: Evaluating Personalized Guidance with Memory Alignment in Lifelong ConversationsarXiv:2609.09664cs.AIHyojeong Yu, Hyukhun Koh, Minsung Kim +211 Sept 2026
Safe to Stop? Risk-Constrained Stopping for Sequential Clinical Diagnosis AgentsarXiv:2609.09678cs.AIYuexin Wu, Vasile Rus11 Sept 2026
Decision Shifts, Lost Label Functionality, and an Inconclusive Grounding Audit in Correctness-Gated Multi-Teacher DistillationarXiv:2609.09702cs.AIXiaofei Feng11 Sept 2026
Which Tokens Should SFT Actually Learn? A Token-Trimming Perspective on Mathematical ReasoningarXiv:2609.09707cs.AIYaning Jia, Chunhui Zhang, Wenxuan Xu +211 Sept 2026
Can Artificial Intelligence Support Healthcare and Mental Health Through Early Cyberbullying Detection ? The Impact of Emotion-Aware AI on Proactive Online SafetyarXiv:2609.09735cs.AIHamed Jelodar, Amir Firouzi, Yen-Wu Lo +211 Sept 2026
LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal AgentsarXiv:2609.09754cs.AIYujin Zhou, Mingxuan Zheng, Chuxue Cao +211 Sept 2026
Procedural Memory Under Change: Reuse and Interference in Controlled Web TasksarXiv:2609.09774cs.AIYanze Cao11 Sept 2026
Proof-Carrying Cognition: Closing the Verification Gap with Reality-Settled RewardarXiv:2609.09776cs.AIEshwar Reddy M, Sourav Karmakar11 Sept 2026
UnitBoost: Managing Compound LLM Systems with a Merge Operator, Not a ModelarXiv:2609.09815cs.AIXing Zhang, Guanghui Wang, Yanwei Cui +211 Sept 2026
The Era by Eon Benchmark: A Generated Enterprise Estate with Exact Ground Truth for Benchmarking LLM AgentsarXiv:2609.09853cs.AIBenjamin Gruenbaum, Doron Porat, Assaf Natanzon +211 Sept 2026
Shifting Relational Paradigms for Affective Computing: Affective Resonance, Vitality Affects, and Vocal Interaction FieldsarXiv:2609.09864cs.AICy Gorman, Yihang Yao11 Sept 2026
AgentAudit: An Open, Extensible Framework for Full-Lifecycle Trust Evaluation of AI AgentsarXiv:2609.09875cs.AIShrey Nag, Sachita, Abhishek Kumar Singh +211 Sept 2026
Scored vs. Generated Readouts in Behavioral Language Models: An Empirical Study of Elicitation FormatarXiv:2609.09882cs.AITouchapon Kraisingkorn, Krittin Pachtrachai, Wachiravit Modecrua11 Sept 2026
Decision Transformer for UAV-Mounted RIS-Assisted Dynamic D2D CommunicationsarXiv:2609.09885cs.AIYaxuan Liu11 Sept 2026
Grounded Evaluation and Repair for NL-to-PDDL Problem GenerationarXiv:2609.09898cs.AIJoana Rosa, Pedro Santos, Valdemar Oliveira +211 Sept 2026
Time-Frequency Geometric Cross-Attention for Chunked Vision-Language-Action ModelsarXiv:2609.09925cs.AIShengye Dong, Haochen Niu, Hao Liu +211 Sept 2026
Structural Process Supervision for Latent Chain-of-Thought ReasoningarXiv:2609.09928cs.AIYiqi Li, Xu Chen, Chen Ju +211 Sept 2026
Belief-State Engine: Augmenting LLMs for Principled Planning Under Partial ObservabilityarXiv:2609.10036cs.AIArnab Chattopadhayay, Debdipta Halder11 Sept 2026
OntologyAligner: Ontology-Aligned Retrieval and Hierarchy-Guided Large Language Model Reranking for Biomedical Ontology NormalizationarXiv:2609.10055cs.AIJie Song, Zhichuan Xu, Ziyu Lu +211 Sept 2026
Reference-Based Bias Detection in LLMs via Relative Representations of Hidden StatesarXiv:2609.10060cs.AIMarek Jeli\'nski, Jan Dubi\'nski, Maciej Chrabaszcz +111 Sept 2026
RAP: Research Attention Prediction Reveals Target-Conditioned Evidence Acquisition BiasesarXiv:2609.10092cs.AIYingqian Wu, Jingcong Liang, Siyuan Wang +211 Sept 2026
Agent-Based ML-LLM Fusion with Self-Optimizing Prompts for Plateau Weather AlertsarXiv:2609.10135cs.AIShuai Yan, Yang Xu, Shan He11 Sept 2026
Kernel-Managed Shared Memory for System-Wide PersonalizationarXiv:2609.10144cs.AIRyan Lum, Yongfeng Zhang11 Sept 2026
Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context LearningarXiv:2609.10177cs.AIMingbo Yang, Wenqiang Wang, Zhaolu Kang +211 Sept 2026
What Should an Agent Forget? Separating What Is Stored from What Is UsedarXiv:2609.10263cs.AIYuhang Li, Yuchen Li11 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.