| Recognizing Is Not Reversing: A Controlled Inversion Test of Fact-Preserving News FramingarXiv:2609.11769 | Yi Liu | 11 Sept 2026 | cs.CL | — | 89 |
| The widening evaluation gap in medical large language model research 2023 to 2026arXiv:2609.11770 | Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif | 11 Sept 2026 | cs.CL | — | 89 |
| Beyond Word Error Rate: A Switch Aware Evaluation of ASR and Audio Language Models on English Yoruba Code-Switched SpeecharXiv:2609.11786 | Chibuzor Okocha, Christan Earl Grant | 11 Sept 2026 | cs.CL | — | 89 |
| Target leakage, not model class, explains reported accuracy in survey-based cardiovascular screening: a leakage-tiered audit of glass-box and tabular foundation modelsarXiv:2609.11838 | Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif +2 | 11 Sept 2026 | cs.CL | — | 89 |
| IndicTriMix: Developing Language Identification Datasets and Models for Tri-Language Code-MixingarXiv:2609.11851 | Pruthwik Mishra, Rudra Trivedi, Avi Patel +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Epistemic orientation predicts legislative effectiveness among members of the US CongressarXiv:2609.11865 | Segun Aroyehun, Stephan Lewandowsky, David Garcia | 11 Sept 2026 | cs.CL | — | 89 |
| Augustinian BabyLM: What Ostensive Definition Can and Cannot Teach a Small Language ModelarXiv:2609.11870 | Lisa Bylinina | 11 Sept 2026 | cs.CL | — | 89 |
| Nuha-Speech: Building General-Purpose Arabic Speech-LLMsarXiv:2609.11892 | Yingzhi Wang, Reem Alhazzani, Muhammad Alqurishi | 11 Sept 2026 | cs.CL | — | 89 |
| Distance generalization in transformers: why bother with positional encoding?arXiv:2609.11913 | Daniel Henrik Nevermann, Claudius Gros | 11 Sept 2026 | cs.CL | — | 89 |
| More than half of recent astronomy papers are written with language-model assistancearXiv:2609.10664 | Serat M. Saad, Yuan-Sen Ting | 11 Sept 2026 | astro-ph.IM | — | 89 |
| BodyCam-VQA: Enhanced Body-Worn Camera Video Captioning via Multimodal Reasoning and Probe Question GenerationarXiv:2609.10815 | Karish Gupta, Matthew Alex, Alex Li +2 | 11 Sept 2026 | cs.CV | — | 89 |
| KuaiRP Series Role-playing Models Technical ReportarXiv:2609.11127 | Yipeng Wang, Ziwei Zhang, Jiahui Zhang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Same Day, Same Story; One Day Ahead, a Different Signal: The Dual Validity of Financial SentimentarXiv:2609.11144 | AS Aravinthkakshan, Laven Srivastava, Harsh Nandwani | 11 Sept 2026 | cs.AI | — | 89 |
| (Whose defaults?) Is artificial intelligence reorienting archaeological methods?arXiv:2609.11198 | Lorenzo Cardarelli, Roberto Ragno | 11 Sept 2026 | cs.CY | — | 89 |
| A Voice-Interactive Multi-Agent System for Smart Operating Rooms: Architecture Design and Key TechnologiesarXiv:2609.11231 | Tianxiang Zhou | 11 Sept 2026 | cs.AI | — | 89 |
| INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry ArchivesarXiv:2609.11261 | Daniel Akselrad, Robert N. Proctor | 11 Sept 2026 | cs.DL | — | 89 |
| Xiaomi-CocktailASR-1 Technical ReportarXiv:2609.11274 | Yiru Zhang, Hang Su, Lichun Fan +2 | 11 Sept 2026 | cs.SD | — | 89 |
| The Semantic Elevation Operator and the Closure of the Undecidable Class under PreservationarXiv:2609.11326 | Jose Pascual Gumbau Mezquita | 11 Sept 2026 | cs.LO | — | 89 |
| Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural UnderstandingarXiv:2609.11772 | Michael Picheny | 11 Sept 2026 | eess.AS | — | 89 |
| SpecGuard: Inference-Time Backdoor Detection For FreearXiv:2609.11799 | Rui Wen, Ahmed Salem, Andrew Paverd +2 | 11 Sept 2026 | cs.CR | — | 89 |
| RetroThinker: Enabling Retrospective Thinking in Speech LLMsarXiv:2609.11864 | Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed +1 | 11 Sept 2026 | eess.AS | — | 89 |
| Biology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR ScreensarXiv:2609.11877 | Carl Edwards, Edward De Brouwer, Xiner Li +2 | 11 Sept 2026 | q-bio.QM | — | 89 |
| MindTopo: Can Foundation Models Reason in Topological Space?arXiv:2609.11900 | Yunfei Ge, Anbang Liu, Qineng Wang +2 | 11 Sept 2026 | cs.AI | — | 89 |
| A Short Survey of Viewing Large Language Models in Legal AspectarXiv:2303.09136 | Zhongxiang Sun | 11 Sept 2026 | cs.CL | — | 89 |
| "Mirror" Large Language Model Evaluations of Depression are Criterion ContaminatedarXiv:2508.05830 | Tong Li, Rasiq Hussain, Mehak Gupta +1 | 11 Sept 2026 | cs.CL | — | 89 |
| The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM SocietiesarXiv:2509.18052 | Jiaxu Zhou, Jen-tse Huang, Xuhui Zhou +2 | 11 Sept 2026 | cs.CL | — | 89 |
| CHRONOBERG: Capturing Language Evolution and Temporal Awareness in Foundation ModelsarXiv:2509.22360 | Niharika Hegde, Subarnaduti Paul, Lars Joel-Frey +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech DetectionarXiv:2510.15685 | Joshua Wolfe Brook, Ilia Markov | 11 Sept 2026 | cs.CL | — | 89 |
| Do Vision-Language Models Understand Visual Persuasiveness? A Diagnosis via Visual Persuasive FactorsarXiv:2511.17036 | Gyuwon Park, Hyounghun Kim | 11 Sept 2026 | cs.CL | — | 89 |
| DeepResearch Bench II: Diagnosing Deep Research Agents via Rubrics from Expert ReportsarXiv:2601.08536 | Ruizhe Li, Mingxuan Du, Benfeng Xu +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Towards Reliable Medical LLMs: Benchmarking and Enhancing Confidence Estimation of Large Language Models in Medical ConsultationarXiv:2601.15645 | Zhiyao Ren, Yibing Zhan, Siyuan Liang +2 | 11 Sept 2026 | cs.CL | — | 89 |
| What Language is This? Ask Your TokenizerarXiv:2602.17655 | Clara Meister, Ahmetcan Yavuz, Pietro Lesci +1 | 11 Sept 2026 | cs.CL | — | 89 |
| Probing for Knowledge Attribution in Large Language ModelsarXiv:2602.22787 | Ivo Brink, Alexander Boer, Dennis Ulmer | 11 Sept 2026 | cs.CL | — | 89 |
| Streaming Translation and Transcription Through Speech-to-Text Causal AlignmentarXiv:2603.11578 | Roman Koshkin, Jeon Haesung, Lianbo Liu +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social InteractionarXiv:2603.17094 | Ryo Kamoi, Ameya Godbole, Binglin Zhou +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Alignment Reduces Expressed but Not Encoded Gender Bias: A Unified Framework and StudyarXiv:2603.24125 | Nour Bouchouchi, Thibault Laugel, Xavier Renard +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Analyzing LLM Reasoning to Uncover Mental Health StigmaarXiv:2604.25053 | Sreehari Sankar, Aliakbar Nafar, Mona Barman +2 | 11 Sept 2026 | cs.CL | — | 89 |
| Timing is Everything: Temporal Scaffolding of Semantic Surprise in HumorarXiv:2605.00143 | Yuxi Ma, Yongqian Peng, Junchen Lyu +2 | 11 Sept 2026 | cs.CL | — | 89 |
| A Recipe for Long-Context Reasoning in Large Language Models via On-Policy Optimization and DistillationarXiv:2605.12227 | Miguel Moura Ramos, Duarte M. Alves, Andr\'e F. T. Martins | 11 Sept 2026 | cs.CL | — | 89 |
| Cross-lingual brain-language model alignment is robust but challenges hierarchical and computational accountsarXiv:2605.21049 | Ni Yang, Rui He, Philipp Homan +2 | 11 Sept 2026 | cs.CL | — | 89 |