| Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary ModelsarXiv:2609.10996 | Yu-Chung Hsiao | 11 Sept 2026 | cs.CL | — | 89 |
| K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language ModelsarXiv:2609.11020 | Yu Sun, Mengyin Lu, Cong Feng +2 | 11 Sept 2026 | cs.CL | — | 89 |
| ProMediConv: Benchmarking Proactive Conversational Agents in Legal Dispute MediationarXiv:2609.11101 | Zesheng Wei, Mengfan Li, Wenhao Liu +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Overview of the NLPCC 2026 Shared Task 11: Agent-Based Experiment Reproduction from Scientific PapersarXiv:2609.11117 | Hanhua Hong, Yizhi Li, Luu Gia Huy +2 | 11 Sept 2026 | cs.CL | — | 89 |
| From Repetition to Recognition: Inductive Discovery of Disinformation NarrativesarXiv:2609.11128 | Max Upravitelev, Veronika Solopova, Jing Yang +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Rubric-Aligned Disentangled Evaluation of Human Simultaneous InterpretingarXiv:2609.11131 | Ziyu Zhang, Satoshi Nakamura | 11 Sept 2026 | cs.CL | — | 89 |
| Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema NormalizationarXiv:2609.11141 | Dong-Jae Koh, Huisu Kim, SeongHwan Yoon +2 | 11 Sept 2026 | cs.CL | — | 89 |
| FlexComp: One Model for Every Ratio in Context CompressionarXiv:2609.11192 | Kaiyan Zhao, Zhongtao Miao, Akiko Aizawa +1 | 11 Sept 2026 | cs.CL | — | 89 |
| Automated Identification of Competing Narratives in Political Discourse on Social MediaarXiv:2609.11202 | Sergej Wildemann, Erick Elejalde | 11 Sept 2026 | cs.CL | — | 89 |
| OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language ModelsarXiv:2609.11244 | Jianjiang Yang, Peihang Li, Shanqing Xu +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Assessing the Reusability of Public Speech Resources for Low-Resource Languages: A Central Kurdish Case StudyarXiv:2609.11246 | Hiwa Asadpour | 11 Sept 2026 | cs.CL | — | 89 |
| The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based MethodsarXiv:2609.11247 | Ioanna Kaffeza, Efthymios Georgiou, Alexandros Potamianos | 11 Sept 2026 | cs.CL | — | 89 |
| Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper AdaptationarXiv:2609.11302 | Maria Frangiadaki, Dimitrios Damianos, Kosmas Kritsis +1 | 11 Sept 2026 | cs.CL | — | 89 |
| MultiHuSE: A Multimodal Dataset for Humour Styles and EmotionsarXiv:2609.11322 | Mary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat | 11 Sept 2026 | cs.CL | — | 89 |
| On the Impact of Anonymization on the Performance of Large Language ModelsarXiv:2609.11335 | Tobias Deu{\ss}er, Max Hahnb\"uck, Lorenz Sparrenberg +2 | 11 Sept 2026 | cs.CL | — | 89 |
| SEAR: Segment-Evidence-Aware Routing for Weak-to-Strong Multilingual Speech MCQarXiv:2609.11355 | Huy Hoang Le, Long-Bao Nguyen, Minh Tri Dao | 11 Sept 2026 | cs.CL | — | 89 |
| TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model OutputsarXiv:2609.11399 | Shenbin Qian, Yves Scherrer | 11 Sept 2026 | cs.CL | — | 89 |
| SWRouter: Similarity-Contractive Window Routing for Multi-Turn Large Language Model ConversationsarXiv:2609.11414 | Yu Wang, Yuchen Li, Rui Kong +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Cross-Lingual Clinical Annotation Projection as Constrained Text Generation: A Six-Language StudyarXiv:2609.11450 | \'Alvaro Rey-Blanes, Francisco J. Moreno-Barea, Francisco J. Veredas | 11 Sept 2026 | cs.CL | — | 89 |
| ReGround: Grounding Reviewer Comments in Multimodal EvidencearXiv:2609.11460 | Serwar Basch, Lizhen Qu, Iryna Gurevych | 11 Sept 2026 | cs.CL | — | 89 |
| Complex-Text Robustness Evaluation and Failure Diagnosis for Low-Resource Multilingual Text-to-SpeecharXiv:2609.11545 | Tianlun Zuo, Ziyu Zhang, Tingzhi Mao +2 | 11 Sept 2026 | cs.CL | — | 89 |
| A Training-Free, Alignment-Free Approach to Corporate Intelligence: Application to SEC FilingsarXiv:2609.11620 | Jean-Fran\c{c}ois Delpech | 11 Sept 2026 | cs.CL | — | 89 |
| Structured Transforms for Low-Overhead Quantization of Language ModelsarXiv:2609.11687 | Daria Cherniuk, Alexander Rudikov, Boris Kashin +1 | 11 Sept 2026 | cs.CL | — | 89 |
| The Eloquence submission for Task 2 of the Interspeech 2026 MLC-SLM challengearXiv:2609.11724 | Jordi Luque, Lorenzo Concina, Marco Matassoni +2 | 11 Sept 2026 | cs.CL | — | 89 |
| RAG-Safety-Bench: Reliable Evaluation of Retrieval-Augmented LLM SafetyarXiv:2609.11758 | Adithiyan Rajan Indira Saravanan, Kathleen C. Fraser | 11 Sept 2026 | cs.CL | — | 89 |
| Component-Aware Differential Privacy for Federated Multilingual Speech-LLMsarXiv:2609.11762 | Jordi Luque, Fernando L\'opez, Aleix Sant | 11 Sept 2026 | cs.CL | — | 89 |
| Recognizing Is Not Reversing: A Controlled Inversion Test of Fact-Preserving News FramingarXiv:2609.11769 | Yi Liu | 11 Sept 2026 | cs.CL | — | 89 |
| The widening evaluation gap in medical large language model research 2023 to 2026arXiv:2609.11770 | Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif | 11 Sept 2026 | cs.CL | — | 89 |
| Beyond Word Error Rate: A Switch Aware Evaluation of ASR and Audio Language Models on English Yoruba Code-Switched SpeecharXiv:2609.11786 | Chibuzor Okocha, Christan Earl Grant | 11 Sept 2026 | cs.CL | — | 89 |
| Target leakage, not model class, explains reported accuracy in survey-based cardiovascular screening: a leakage-tiered audit of glass-box and tabular foundation modelsarXiv:2609.11838 | Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif +2 | 11 Sept 2026 | cs.CL | — | 89 |
| IndicTriMix: Developing Language Identification Datasets and Models for Tri-Language Code-MixingarXiv:2609.11851 | Pruthwik Mishra, Rudra Trivedi, Avi Patel +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Epistemic orientation predicts legislative effectiveness among members of the US CongressarXiv:2609.11865 | Segun Aroyehun, Stephan Lewandowsky, David Garcia | 11 Sept 2026 | cs.CL | — | 89 |
| Augustinian BabyLM: What Ostensive Definition Can and Cannot Teach a Small Language ModelarXiv:2609.11870 | Lisa Bylinina | 11 Sept 2026 | cs.CL | — | 89 |
| Nuha-Speech: Building General-Purpose Arabic Speech-LLMsarXiv:2609.11892 | Yingzhi Wang, Reem Alhazzani, Muhammad Alqurishi | 11 Sept 2026 | cs.CL | — | 89 |
| Distance generalization in transformers: why bother with positional encoding?arXiv:2609.11913 | Daniel Henrik Nevermann, Claudius Gros | 11 Sept 2026 | cs.CL | — | 89 |
| More than half of recent astronomy papers are written with language-model assistancearXiv:2609.10664 | Serat M. Saad, Yuan-Sen Ting | 11 Sept 2026 | astro-ph.IM | — | 89 |
| BodyCam-VQA: Enhanced Body-Worn Camera Video Captioning via Multimodal Reasoning and Probe Question GenerationarXiv:2609.10815 | Karish Gupta, Matthew Alex, Alex Li +2 | 11 Sept 2026 | cs.CV | — | 89 |
| KuaiRP Series Role-playing Models Technical ReportarXiv:2609.11127 | Yipeng Wang, Ziwei Zhang, Jiahui Zhang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Same Day, Same Story; One Day Ahead, a Different Signal: The Dual Validity of Financial SentimentarXiv:2609.11144 | AS Aravinthkakshan, Laven Srivastava, Harsh Nandwani | 11 Sept 2026 | cs.AI | — | 89 |
| (Whose defaults?) Is artificial intelligence reorienting archaeological methods?arXiv:2609.11198 | Lorenzo Cardarelli, Roberto Ragno | 11 Sept 2026 | cs.CY | — | 89 |