| PRAGMA: Evaluating Personalized Guidance with Memory Alignment in Lifelong ConversationsarXiv:2609.09664 | Hyojeong Yu, Hyukhun Koh, Minsung Kim +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Safe to Stop? Risk-Constrained Stopping for Sequential Clinical Diagnosis AgentsarXiv:2609.09678 | Yuexin Wu, Vasile Rus | 11 Sept 2026 | cs.AI | — | 89 |
| Decision Shifts, Lost Label Functionality, and an Inconclusive Grounding Audit in Correctness-Gated Multi-Teacher DistillationarXiv:2609.09702 | Xiaofei Feng | 11 Sept 2026 | cs.AI | — | 89 |
| Which Tokens Should SFT Actually Learn? A Token-Trimming Perspective on Mathematical ReasoningarXiv:2609.09707 | Yaning Jia, Chunhui Zhang, Wenxuan Xu +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Can Artificial Intelligence Support Healthcare and Mental Health Through Early Cyberbullying Detection ? The Impact of Emotion-Aware AI on Proactive Online SafetyarXiv:2609.09735 | Hamed Jelodar, Amir Firouzi, Yen-Wu Lo +2 | 11 Sept 2026 | cs.AI | — | 89 |
| LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal AgentsarXiv:2609.09754 | Yujin Zhou, Mingxuan Zheng, Chuxue Cao +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Procedural Memory Under Change: Reuse and Interference in Controlled Web TasksarXiv:2609.09774 | Yanze Cao | 11 Sept 2026 | cs.AI | — | 89 |
| Proof-Carrying Cognition: Closing the Verification Gap with Reality-Settled RewardarXiv:2609.09776 | Eshwar Reddy M, Sourav Karmakar | 11 Sept 2026 | cs.AI | — | 89 |
| UnitBoost: Managing Compound LLM Systems with a Merge Operator, Not a ModelarXiv:2609.09815 | Xing Zhang, Guanghui Wang, Yanwei Cui +2 | 11 Sept 2026 | cs.AI | — | 89 |
| The Era by Eon Benchmark: A Generated Enterprise Estate with Exact Ground Truth for Benchmarking LLM AgentsarXiv:2609.09853 | Benjamin Gruenbaum, Doron Porat, Assaf Natanzon +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Shifting Relational Paradigms for Affective Computing: Affective Resonance, Vitality Affects, and Vocal Interaction FieldsarXiv:2609.09864 | Cy Gorman, Yihang Yao | 11 Sept 2026 | cs.AI | — | 89 |
| AgentAudit: An Open, Extensible Framework for Full-Lifecycle Trust Evaluation of AI AgentsarXiv:2609.09875 | Shrey Nag, Sachita, Abhishek Kumar Singh +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Scored vs. Generated Readouts in Behavioral Language Models: An Empirical Study of Elicitation FormatarXiv:2609.09882 | Touchapon Kraisingkorn, Krittin Pachtrachai, Wachiravit Modecrua | 11 Sept 2026 | cs.AI | — | 89 |
| Decision Transformer for UAV-Mounted RIS-Assisted Dynamic D2D CommunicationsarXiv:2609.09885 | Yaxuan Liu | 11 Sept 2026 | cs.AI | — | 89 |
| Grounded Evaluation and Repair for NL-to-PDDL Problem GenerationarXiv:2609.09898 | Joana Rosa, Pedro Santos, Valdemar Oliveira +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Time-Frequency Geometric Cross-Attention for Chunked Vision-Language-Action ModelsarXiv:2609.09925 | Shengye Dong, Haochen Niu, Hao Liu +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Structural Process Supervision for Latent Chain-of-Thought ReasoningarXiv:2609.09928 | Yiqi Li, Xu Chen, Chen Ju +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Belief-State Engine: Augmenting LLMs for Principled Planning Under Partial ObservabilityarXiv:2609.10036 | Arnab Chattopadhayay, Debdipta Halder | 11 Sept 2026 | cs.AI | — | 89 |
| OntologyAligner: Ontology-Aligned Retrieval and Hierarchy-Guided Large Language Model Reranking for Biomedical Ontology NormalizationarXiv:2609.10055 | Jie Song, Zhichuan Xu, Ziyu Lu +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Reference-Based Bias Detection in LLMs via Relative Representations of Hidden StatesarXiv:2609.10060 | Marek Jeli\'nski, Jan Dubi\'nski, Maciej Chrabaszcz +1 | 11 Sept 2026 | cs.AI | — | 89 |
| RAP: Research Attention Prediction Reveals Target-Conditioned Evidence Acquisition BiasesarXiv:2609.10092 | Yingqian Wu, Jingcong Liang, Siyuan Wang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Agent-Based ML-LLM Fusion with Self-Optimizing Prompts for Plateau Weather AlertsarXiv:2609.10135 | Shuai Yan, Yang Xu, Shan He | 11 Sept 2026 | cs.AI | — | 89 |
| Kernel-Managed Shared Memory for System-Wide PersonalizationarXiv:2609.10144 | Ryan Lum, Yongfeng Zhang | 11 Sept 2026 | cs.AI | — | 89 |
| Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context LearningarXiv:2609.10177 | Mingbo Yang, Wenqiang Wang, Zhaolu Kang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Why Sample What You Can Enumerate? Exact Policy Optimization for Genomic Tool SelectionarXiv:2609.10221 | Haoyue Liu, Xiaoyu Ma, Ye Chen +2 | 11 Sept 2026 | cs.AI | — | 89 |
| What Should an Agent Forget? Separating What Is Stored from What Is UsedarXiv:2609.10263 | Yuhang Li, Yuchen Li | 11 Sept 2026 | cs.AI | — | 89 |
| TRACE: Training Reasoning Agents for Causal Exploration with Synthesized RewardsarXiv:2609.10315 | Rui Sun, Zhan Shi, Bing He | 11 Sept 2026 | cs.AI | — | 89 |
| From Symbolic Perception to Logical Deduction: A Framework for Guiding Language Models in Geometric ReasoningarXiv:2609.10335 | Weichen Dai, Rafael Medeiros Cabral, Ziyi Shou +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Cyber-Financial Contagion: Modeling the Propagation of an AI Vendor Compromise Through the Banking SystemarXiv:2609.10350 | Alex Leytes | 11 Sept 2026 | cs.AI | — | 89 |
| Fortunate Recall: Ontology-Driven Memory Lifecycle Management for Persistent Coherence in LLMsarXiv:2609.10413 | Ansuman Mullick, Eray T\"uz\"un | 11 Sept 2026 | cs.AI | — | 89 |
| ConvMem: Convolutional Memory for Long-Context ReasoningarXiv:2609.10441 | Hongming Zhang, Zhaozhen Gu, Fengshuo Bai +2 | 11 Sept 2026 | cs.AI | — | 89 |
| JarvisGUI: Towards Cross-Device GUI Agents with Dynamic Task CompositionarXiv:2609.10451 | Zixiang Chen, Yuheng Lu, Zihao Cheng +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Quantifying Logical Consistency in Transformers via Query-Key AlignmentarXiv:2502.17017 | Eduard Tulchinskii, Anastasia Voznyuk, Laida Kushnareva +2 | 11 Sept 2026 | cs.CL | — | 89 |
| From Plausible to Actionable: A Position on LLM Self-ExplanationsarXiv:2607.15957 | Elize Herrewijnen, Benedetta Muscato, Gizem Gezici +1 | 11 Sept 2026 | cs.CL | — | 89 |
| Characterizing Text Branch Sensitivity in Medical Vision-Language Segmentation via Evidence DecouplingarXiv:2609.02663 | Ziquan Liu, Zhewei Zhu, Xuyang Shi | 11 Sept 2026 | cs.CV | — | 89 |
| Trust Me, I'm Your Developer: Self-Issued Authentication in Large Language ModelsarXiv:2609.03247 | Syed Ghazanfar Abbas, Dongyan Xu | 11 Sept 2026 | cs.CR | — | 89 |
| AgenticGen: Reward-Guided Agentic Video Generation for AdvertisingarXiv:2609.09187 | Xingyuan Bu, Chengru Song, Hao Zhou +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Reliability-Aware Hybrid-K Ensemble Selection for Cervical Cytology Classification: Integrating Discrimination, Calibration, and Selective PredictionarXiv:2609.09189 | Nisreen Albzour, Sarah S. Lam | 11 Sept 2026 | eess.IV | — | 89 |
| AgentHijack: Visual Patch Attacks on Multimodal Computer-Use AgentsarXiv:2609.09212 | Zhihao Liu, Hongyu Sun, Zhiyuan Fu +2 | 11 Sept 2026 | cs.CR | — | 89 |
| Geometry Conditioning in an Embodied SLM: Training Controls and Robustness Diagnostics in a 0.8B Hybrid ModelarXiv:2609.09213 | Hao Li, Haofei Sun, Lin He | 11 Sept 2026 | cs.RO | — | 89 |