Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages.

210 total

Reset
Papers
TitleAuthorsPublishedCategoriesOrganization / venueQuality
Recognizing Is Not Reversing: A Controlled Inversion Test of Fact-Preserving News FramingarXiv:2609.11769Yi Liu11 Sept 2026cs.CL89
The widening evaluation gap in medical large language model research 2023 to 2026arXiv:2609.11770Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif11 Sept 2026cs.CL89
Beyond Word Error Rate: A Switch Aware Evaluation of ASR and Audio Language Models on English Yoruba Code-Switched SpeecharXiv:2609.11786Chibuzor Okocha, Christan Earl Grant11 Sept 2026cs.CL89
Target leakage, not model class, explains reported accuracy in survey-based cardiovascular screening: a leakage-tiered audit of glass-box and tabular foundation modelsarXiv:2609.11838Raad Bin Tareaf, Murad Al-Rajab, Samia Loucif +211 Sept 2026cs.CL89
IndicTriMix: Developing Language Identification Datasets and Models for Tri-Language Code-MixingarXiv:2609.11851Pruthwik Mishra, Rudra Trivedi, Avi Patel +211 Sept 2026cs.CL89
Epistemic orientation predicts legislative effectiveness among members of the US CongressarXiv:2609.11865Segun Aroyehun, Stephan Lewandowsky, David Garcia11 Sept 2026cs.CL89
Augustinian BabyLM: What Ostensive Definition Can and Cannot Teach a Small Language ModelarXiv:2609.11870Lisa Bylinina11 Sept 2026cs.CL89
Nuha-Speech: Building General-Purpose Arabic Speech-LLMsarXiv:2609.11892Yingzhi Wang, Reem Alhazzani, Muhammad Alqurishi11 Sept 2026cs.CL89
Distance generalization in transformers: why bother with positional encoding?arXiv:2609.11913Daniel Henrik Nevermann, Claudius Gros11 Sept 2026cs.CL89
More than half of recent astronomy papers are written with language-model assistancearXiv:2609.10664Serat M. Saad, Yuan-Sen Ting11 Sept 2026astro-ph.IM89
BodyCam-VQA: Enhanced Body-Worn Camera Video Captioning via Multimodal Reasoning and Probe Question GenerationarXiv:2609.10815Karish Gupta, Matthew Alex, Alex Li +211 Sept 2026cs.CV89
KuaiRP Series Role-playing Models Technical ReportarXiv:2609.11127Yipeng Wang, Ziwei Zhang, Jiahui Zhang +211 Sept 2026cs.AI89
Same Day, Same Story; One Day Ahead, a Different Signal: The Dual Validity of Financial SentimentarXiv:2609.11144AS Aravinthkakshan, Laven Srivastava, Harsh Nandwani11 Sept 2026cs.AI89
(Whose defaults?) Is artificial intelligence reorienting archaeological methods?arXiv:2609.11198Lorenzo Cardarelli, Roberto Ragno11 Sept 2026cs.CY89
A Voice-Interactive Multi-Agent System for Smart Operating Rooms: Architecture Design and Key TechnologiesarXiv:2609.11231Tianxiang Zhou11 Sept 2026cs.AI89
INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry ArchivesarXiv:2609.11261Daniel Akselrad, Robert N. Proctor11 Sept 2026cs.DL89
Xiaomi-CocktailASR-1 Technical ReportarXiv:2609.11274Yiru Zhang, Hang Su, Lichun Fan +211 Sept 2026cs.SD89
The Semantic Elevation Operator and the Closure of the Undecidable Class under PreservationarXiv:2609.11326Jose Pascual Gumbau Mezquita11 Sept 2026cs.LO89
Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural UnderstandingarXiv:2609.11772Michael Picheny11 Sept 2026eess.AS89
SpecGuard: Inference-Time Backdoor Detection For FreearXiv:2609.11799Rui Wen, Ahmed Salem, Andrew Paverd +211 Sept 2026cs.CR89
RetroThinker: Enabling Retrospective Thinking in Speech LLMsarXiv:2609.11864Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed +111 Sept 2026eess.AS89
Biology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR ScreensarXiv:2609.11877Carl Edwards, Edward De Brouwer, Xiner Li +211 Sept 2026q-bio.QM89
MindTopo: Can Foundation Models Reason in Topological Space?arXiv:2609.11900Yunfei Ge, Anbang Liu, Qineng Wang +211 Sept 2026cs.AI89
A Short Survey of Viewing Large Language Models in Legal AspectarXiv:2303.09136Zhongxiang Sun11 Sept 2026cs.CL89
"Mirror" Large Language Model Evaluations of Depression are Criterion ContaminatedarXiv:2508.05830Tong Li, Rasiq Hussain, Mehak Gupta +111 Sept 2026cs.CL89
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM SocietiesarXiv:2509.18052Jiaxu Zhou, Jen-tse Huang, Xuhui Zhou +211 Sept 2026cs.CL89
CHRONOBERG: Capturing Language Evolution and Temporal Awareness in Foundation ModelsarXiv:2509.22360Niharika Hegde, Subarnaduti Paul, Lars Joel-Frey +211 Sept 2026cs.CL89
Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech DetectionarXiv:2510.15685Joshua Wolfe Brook, Ilia Markov11 Sept 2026cs.CL89
Do Vision-Language Models Understand Visual Persuasiveness? A Diagnosis via Visual Persuasive FactorsarXiv:2511.17036Gyuwon Park, Hyounghun Kim11 Sept 2026cs.CL89
DeepResearch Bench II: Diagnosing Deep Research Agents via Rubrics from Expert ReportsarXiv:2601.08536Ruizhe Li, Mingxuan Du, Benfeng Xu +211 Sept 2026cs.CL89
Towards Reliable Medical LLMs: Benchmarking and Enhancing Confidence Estimation of Large Language Models in Medical ConsultationarXiv:2601.15645Zhiyao Ren, Yibing Zhan, Siyuan Liang +211 Sept 2026cs.CL89
What Language is This? Ask Your TokenizerarXiv:2602.17655Clara Meister, Ahmetcan Yavuz, Pietro Lesci +111 Sept 2026cs.CL89
Probing for Knowledge Attribution in Large Language ModelsarXiv:2602.22787Ivo Brink, Alexander Boer, Dennis Ulmer11 Sept 2026cs.CL89
Streaming Translation and Transcription Through Speech-to-Text Causal AlignmentarXiv:2603.11578Roman Koshkin, Jeon Haesung, Lianbo Liu +211 Sept 2026cs.CL89
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social InteractionarXiv:2603.17094Ryo Kamoi, Ameya Godbole, Binglin Zhou +211 Sept 2026cs.CL89
Alignment Reduces Expressed but Not Encoded Gender Bias: A Unified Framework and StudyarXiv:2603.24125Nour Bouchouchi, Thibault Laugel, Xavier Renard +211 Sept 2026cs.CL89
Analyzing LLM Reasoning to Uncover Mental Health StigmaarXiv:2604.25053Sreehari Sankar, Aliakbar Nafar, Mona Barman +211 Sept 2026cs.CL89
Timing is Everything: Temporal Scaffolding of Semantic Surprise in HumorarXiv:2605.00143Yuxi Ma, Yongqian Peng, Junchen Lyu +211 Sept 2026cs.CL89
A Recipe for Long-Context Reasoning in Large Language Models via On-Policy Optimization and DistillationarXiv:2605.12227Miguel Moura Ramos, Duarte M. Alves, Andr\'e F. T. Martins11 Sept 2026cs.CL89
Cross-lingual brain-language model alignment is robust but challenges hierarchical and computational accountsarXiv:2605.21049Ni Yang, Rui He, Philipp Homan +211 Sept 2026cs.CL89

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Each paper page shows the abstract, related models and every source snapshot.