| Reason Through the Latent! Making Latent Visual Reasoning NecessaryarXiv:2609.06746 | Suhyeong Park, Junha Jung, Jaewoo Kang | 11 Sept 2026 | cs.AI | — | 89 |
| Limitations of Automated Simulatability: LLM Simulators Can Bypass ExplanationsarXiv:2609.08585 | Antonin Poch\'e, Fanny Jourdan, Nils Feldhus +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Data-Efficient Language Modeling: From Frontier Advancement to Principle-Guided Model ImprovementarXiv:2609.10702 | Shuxing Yang, Kaihao Zhu, Junjie Yang +2 | 11 Sept 2026 | cs.CL | — | 89 |
| NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept PredictionarXiv:2609.10715 | NCP Team, Jiaqi Cao, Chiyu Chen +2 | 11 Sept 2026 | cs.CL | — | 84 |
| CMNIE: An Information Extraction Benchmark for Chinese Military NewsarXiv:2609.10722 | Yan Yu, Mengna Zhu, Zhenyu Song +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity LinkingarXiv:2609.10745 | Parinthapat Pengpun, Simran Khanuja, Graham Neubig | 11 Sept 2026 | cs.CL | — | 81 |
| Analyzing Traditional and Neural Approaches to Multilingual Readability AssessmentarXiv:2609.10792 | Joshua Wong, Chris Tanner | 11 Sept 2026 | cs.CL | — | 89 |
| Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error CorrectionarXiv:2609.10810 | Kateryna Karpo, Artem Chernodub | 11 Sept 2026 | cs.CL | — | 89 |
| Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current ArchitecturesarXiv:2609.10893 | Victor Mazzotti, Luiz Pereira, Marina Bitencourt dos Santos +2 | 11 Sept 2026 | cs.CL | — | 89 |
| LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease DetectionarXiv:2609.10896 | Xiao Wei, Yuqin Lin, Yaru Cao +2 | 11 Sept 2026 | cs.CL | — | 89 |
| SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query GraphsarXiv:2609.10901 | Jiacheng Sang, Mengyuan Li, Sanxing Chen +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender SystemarXiv:2609.10922 | Ming Li, Dai Li, Xuying Ning +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Using Semantic Uncertainty to Estimate Transition Relevance in Turn-takingarXiv:2609.10934 | Muhammad Umair, Jan P. de Ruiter | 11 Sept 2026 | cs.CL | — | 89 |
| Distribution-aware Language Neuron Identification in Multilingual Large Language ModelsarXiv:2609.10993 | Minjun Kim, Inho Won, Junghun Yuk +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary ModelsarXiv:2609.10996 | Yu-Chung Hsiao | 11 Sept 2026 | cs.CL | — | 89 |
| K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language ModelsarXiv:2609.11020 | Yu Sun, Mengyin Lu, Cong Feng +2 | 11 Sept 2026 | cs.CL | — | 89 |
| ProMediConv: Benchmarking Proactive Conversational Agents in Legal Dispute MediationarXiv:2609.11101 | Zesheng Wei, Mengfan Li, Wenhao Liu +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Overview of the NLPCC 2026 Shared Task 11: Agent-Based Experiment Reproduction from Scientific PapersarXiv:2609.11117 | Hanhua Hong, Yizhi Li, Luu Gia Huy +2 | 11 Sept 2026 | cs.CL | — | 89 |
| From Repetition to Recognition: Inductive Discovery of Disinformation NarrativesarXiv:2609.11128 | Max Upravitelev, Veronika Solopova, Jing Yang +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Rubric-Aligned Disentangled Evaluation of Human Simultaneous InterpretingarXiv:2609.11131 | Ziyu Zhang, Satoshi Nakamura | 11 Sept 2026 | cs.CL | — | 89 |
| Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema NormalizationarXiv:2609.11141 | Dong-Jae Koh, Huisu Kim, SeongHwan Yoon +2 | 11 Sept 2026 | cs.CL | — | 89 |
| FlexComp: One Model for Every Ratio in Context CompressionarXiv:2609.11192 | Kaiyan Zhao, Zhongtao Miao, Akiko Aizawa +1 | 11 Sept 2026 | cs.CL | — | 89 |
| Automated Identification of Competing Narratives in Political Discourse on Social MediaarXiv:2609.11202 | Sergej Wildemann, Erick Elejalde | 11 Sept 2026 | cs.CL | — | 89 |
| OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language ModelsarXiv:2609.11244 | Jianjiang Yang, Peihang Li, Shanqing Xu +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Assessing the Reusability of Public Speech Resources for Low-Resource Languages: A Central Kurdish Case StudyarXiv:2609.11246 | Hiwa Asadpour | 11 Sept 2026 | cs.CL | — | 89 |
| The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based MethodsarXiv:2609.11247 | Ioanna Kaffeza, Efthymios Georgiou, Alexandros Potamianos | 11 Sept 2026 | cs.CL | — | 89 |
| Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper AdaptationarXiv:2609.11302 | Maria Frangiadaki, Dimitrios Damianos, Kosmas Kritsis +1 | 11 Sept 2026 | cs.CL | — | 89 |
| MultiHuSE: A Multimodal Dataset for Humour Styles and EmotionsarXiv:2609.11322 | Mary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat | 11 Sept 2026 | cs.CL | — | 89 |
| On the Impact of Anonymization on the Performance of Large Language ModelsarXiv:2609.11335 | Tobias Deu{\ss}er, Max Hahnb\"uck, Lorenz Sparrenberg +2 | 11 Sept 2026 | cs.CL | — | 89 |
| SEAR: Segment-Evidence-Aware Routing for Weak-to-Strong Multilingual Speech MCQarXiv:2609.11355 | Huy Hoang Le, Long-Bao Nguyen, Minh Tri Dao | 11 Sept 2026 | cs.CL | — | 89 |
| TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model OutputsarXiv:2609.11399 | Shenbin Qian, Yves Scherrer | 11 Sept 2026 | cs.CL | — | 89 |
| SWRouter: Similarity-Contractive Window Routing for Multi-Turn Large Language Model ConversationsarXiv:2609.11414 | Yu Wang, Yuchen Li, Rui Kong +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Cross-Lingual Clinical Annotation Projection as Constrained Text Generation: A Six-Language StudyarXiv:2609.11450 | \'Alvaro Rey-Blanes, Francisco J. Moreno-Barea, Francisco J. Veredas | 11 Sept 2026 | cs.CL | — | 89 |
| ReGround: Grounding Reviewer Comments in Multimodal EvidencearXiv:2609.11460 | Serwar Basch, Lizhen Qu, Iryna Gurevych | 11 Sept 2026 | cs.CL | — | 89 |
| Complex-Text Robustness Evaluation and Failure Diagnosis for Low-Resource Multilingual Text-to-SpeecharXiv:2609.11545 | Tianlun Zuo, Ziyu Zhang, Tingzhi Mao +2 | 11 Sept 2026 | cs.CL | — | 89 |
| A Training-Free, Alignment-Free Approach to Corporate Intelligence: Application to SEC FilingsarXiv:2609.11620 | Jean-Fran\c{c}ois Delpech | 11 Sept 2026 | cs.CL | — | 89 |
| Structured Transforms for Low-Overhead Quantization of Language ModelsarXiv:2609.11687 | Daria Cherniuk, Alexander Rudikov, Boris Kashin +1 | 11 Sept 2026 | cs.CL | — | 89 |
| The Eloquence submission for Task 2 of the Interspeech 2026 MLC-SLM challengearXiv:2609.11724 | Jordi Luque, Lorenzo Concina, Marco Matassoni +2 | 11 Sept 2026 | cs.CL | — | 89 |
| RAG-Safety-Bench: Reliable Evaluation of Retrieval-Augmented LLM SafetyarXiv:2609.11758 | Adithiyan Rajan Indira Saravanan, Kathleen C. Fraser | 11 Sept 2026 | cs.CL | — | 89 |
| Component-Aware Differential Privacy for Federated Multilingual Speech-LLMsarXiv:2609.11762 | Jordi Luque, Fernando L\'opez, Aleix Sant | 11 Sept 2026 | cs.CL | — | 89 |