| Valerant: An Automatic Navigable Game Map Generator via Action-Conditioned World Model ExplorationarXiv:2609.09418 | Yiran Qiao, Feng Wang, Jing Ma | 11 Sept 2026 | cs.AI | — | 89 |
| XAI-Arena: Can LLMs Assess the Quality of XAI Explanations?arXiv:2609.09428 | Yanfei Hu Fleischhauer, Alona Zharova, Nadja Klein +1 | 11 Sept 2026 | cs.AI | — | 89 |
| Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal RepresentationsarXiv:2609.09448 | Priyanka Mary Mammen, Emil Joswin, Srujananjali Medicherla | 11 Sept 2026 | cs.AI | — | 89 |
| ContractEval: Query-Conditioned Execution Matching for Procedural Instruction ConformancearXiv:2609.09458 | Praphul Singh, Shanu Kumar, Akshat Agarwal +1 | 11 Sept 2026 | cs.AI | — | 89 |
| Multi-Agent Agentic Graph Learning via Structural SignaturesarXiv:2609.09565 | Liang Qu, Jianxin Li, Hua Wang | 11 Sept 2026 | cs.AI | — | 89 |
| CityPlanner: A Sandbox Agent for Executable Urban PlanningarXiv:2609.09578 | Wentao Zhang, Jingyuan Wang, Zetong Zhou +2 | 11 Sept 2026 | cs.AI | — | 89 |
| A Function-Space Approach to the Statistical Mechanics of Learning DynamicsarXiv:2609.09589 | Yizhou Zhang, Weichen Wu, Lun Du +1 | 11 Sept 2026 | cs.AI | — | 89 |
| From State Synchronization to Cognitive Self-Evolution: An Operational Architecture for Cognitive Digital TwinsarXiv:2609.09625 | Haoran Gao, An Li, Zhen Li +1 | 11 Sept 2026 | cs.AI | — | 89 |
| Seven Sources of Physical AI Capability FormationarXiv:2609.09627 | Gang Chen | 11 Sept 2026 | cs.AI | — | 89 |
| RobustSGPO: Search-Space Control for Agent Harness EvolutionarXiv:2609.09646 | Zibo Zhao, Jijun Shi, Mo Zhou +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Black-Box Red Teaming of Agentic AI: A Taxonomy-Driven Framework for Automated Risk DiscoveryarXiv:2609.09647 | Divyanshu Kumar, Nitin Aravind Birur, Tanay Baswa +2 | 11 Sept 2026 | cs.AI | — | 89 |
| RESCUE-BENCH: Towards Relation-Aware Multi-Party Emotional Support Conversation SystemsarXiv:2609.09657 | Haichuan Hu, Yang Xiao, Mingni Tang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| PRAGMA: Evaluating Personalized Guidance with Memory Alignment in Lifelong ConversationsarXiv:2609.09664 | Hyojeong Yu, Hyukhun Koh, Minsung Kim +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Safe to Stop? Risk-Constrained Stopping for Sequential Clinical Diagnosis AgentsarXiv:2609.09678 | Yuexin Wu, Vasile Rus | 11 Sept 2026 | cs.AI | — | 89 |
| Decision Shifts, Lost Label Functionality, and an Inconclusive Grounding Audit in Correctness-Gated Multi-Teacher DistillationarXiv:2609.09702 | Xiaofei Feng | 11 Sept 2026 | cs.AI | — | 89 |
| Which Tokens Should SFT Actually Learn? A Token-Trimming Perspective on Mathematical ReasoningarXiv:2609.09707 | Yaning Jia, Chunhui Zhang, Wenxuan Xu +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Can Artificial Intelligence Support Healthcare and Mental Health Through Early Cyberbullying Detection ? The Impact of Emotion-Aware AI on Proactive Online SafetyarXiv:2609.09735 | Hamed Jelodar, Amir Firouzi, Yen-Wu Lo +2 | 11 Sept 2026 | cs.AI | — | 89 |
| LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal AgentsarXiv:2609.09754 | Yujin Zhou, Mingxuan Zheng, Chuxue Cao +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Procedural Memory Under Change: Reuse and Interference in Controlled Web TasksarXiv:2609.09774 | Yanze Cao | 11 Sept 2026 | cs.AI | — | 89 |
| Proof-Carrying Cognition: Closing the Verification Gap with Reality-Settled RewardarXiv:2609.09776 | Eshwar Reddy M, Sourav Karmakar | 11 Sept 2026 | cs.AI | — | 89 |
| UnitBoost: Managing Compound LLM Systems with a Merge Operator, Not a ModelarXiv:2609.09815 | Xing Zhang, Guanghui Wang, Yanwei Cui +2 | 11 Sept 2026 | cs.AI | — | 89 |
| The Era by Eon Benchmark: A Generated Enterprise Estate with Exact Ground Truth for Benchmarking LLM AgentsarXiv:2609.09853 | Benjamin Gruenbaum, Doron Porat, Assaf Natanzon +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Shifting Relational Paradigms for Affective Computing: Affective Resonance, Vitality Affects, and Vocal Interaction FieldsarXiv:2609.09864 | Cy Gorman, Yihang Yao | 11 Sept 2026 | cs.AI | — | 89 |
| AgentAudit: An Open, Extensible Framework for Full-Lifecycle Trust Evaluation of AI AgentsarXiv:2609.09875 | Shrey Nag, Sachita, Abhishek Kumar Singh +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Scored vs. Generated Readouts in Behavioral Language Models: An Empirical Study of Elicitation FormatarXiv:2609.09882 | Touchapon Kraisingkorn, Krittin Pachtrachai, Wachiravit Modecrua | 11 Sept 2026 | cs.AI | — | 89 |
| Decision Transformer for UAV-Mounted RIS-Assisted Dynamic D2D CommunicationsarXiv:2609.09885 | Yaxuan Liu | 11 Sept 2026 | cs.AI | — | 89 |
| Grounded Evaluation and Repair for NL-to-PDDL Problem GenerationarXiv:2609.09898 | Joana Rosa, Pedro Santos, Valdemar Oliveira +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Time-Frequency Geometric Cross-Attention for Chunked Vision-Language-Action ModelsarXiv:2609.09925 | Shengye Dong, Haochen Niu, Hao Liu +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Structural Process Supervision for Latent Chain-of-Thought ReasoningarXiv:2609.09928 | Yiqi Li, Xu Chen, Chen Ju +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Belief-State Engine: Augmenting LLMs for Principled Planning Under Partial ObservabilityarXiv:2609.10036 | Arnab Chattopadhayay, Debdipta Halder | 11 Sept 2026 | cs.AI | — | 89 |
| OntologyAligner: Ontology-Aligned Retrieval and Hierarchy-Guided Large Language Model Reranking for Biomedical Ontology NormalizationarXiv:2609.10055 | Jie Song, Zhichuan Xu, Ziyu Lu +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Reference-Based Bias Detection in LLMs via Relative Representations of Hidden StatesarXiv:2609.10060 | Marek Jeli\'nski, Jan Dubi\'nski, Maciej Chrabaszcz +1 | 11 Sept 2026 | cs.AI | — | 89 |
| RAP: Research Attention Prediction Reveals Target-Conditioned Evidence Acquisition BiasesarXiv:2609.10092 | Yingqian Wu, Jingcong Liang, Siyuan Wang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Agent-Based ML-LLM Fusion with Self-Optimizing Prompts for Plateau Weather AlertsarXiv:2609.10135 | Shuai Yan, Yang Xu, Shan He | 11 Sept 2026 | cs.AI | — | 89 |
| Kernel-Managed Shared Memory for System-Wide PersonalizationarXiv:2609.10144 | Ryan Lum, Yongfeng Zhang | 11 Sept 2026 | cs.AI | — | 89 |
| Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context LearningarXiv:2609.10177 | Mingbo Yang, Wenqiang Wang, Zhaolu Kang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Why Sample What You Can Enumerate? Exact Policy Optimization for Genomic Tool SelectionarXiv:2609.10221 | Haoyue Liu, Xiaoyu Ma, Ye Chen +2 | 11 Sept 2026 | cs.AI | — | 89 |
| What Should an Agent Forget? Separating What Is Stored from What Is UsedarXiv:2609.10263 | Yuhang Li, Yuchen Li | 11 Sept 2026 | cs.AI | — | 89 |
| TRACE: Training Reasoning Agents for Causal Exploration with Synthesized RewardsarXiv:2609.10315 | Rui Sun, Zhan Shi, Bing He | 11 Sept 2026 | cs.AI | — | 89 |
| From Symbolic Perception to Logical Deduction: A Framework for Guiding Language Models in Geometric ReasoningarXiv:2609.10335 | Weichen Dai, Rafael Medeiros Cabral, Ziyi Shou +2 | 11 Sept 2026 | cs.AI | — | 89 |