Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages.

1,037 total

Reset
Papers
TitleAuthorsPublishedCategoriesOrganization / venueQuality
A Voice-Interactive Multi-Agent System for Smart Operating Rooms: Architecture Design and Key TechnologiesarXiv:2609.11231Tianxiang Zhou11 Sept 2026cs.AI89
INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry ArchivesarXiv:2609.11261Daniel Akselrad, Robert N. Proctor11 Sept 2026cs.DL89
Xiaomi-CocktailASR-1 Technical ReportarXiv:2609.11274Yiru Zhang, Hang Su, Lichun Fan +211 Sept 2026cs.SD89
The Semantic Elevation Operator and the Closure of the Undecidable Class under PreservationarXiv:2609.11326Jose Pascual Gumbau Mezquita11 Sept 2026cs.LO89
Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural UnderstandingarXiv:2609.11772Michael Picheny11 Sept 2026eess.AS89
SpecGuard: Inference-Time Backdoor Detection For FreearXiv:2609.11799Rui Wen, Ahmed Salem, Andrew Paverd +211 Sept 2026cs.CR89
RetroThinker: Enabling Retrospective Thinking in Speech LLMsarXiv:2609.11864Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed +111 Sept 2026eess.AS89
Biology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR ScreensarXiv:2609.11877Carl Edwards, Edward De Brouwer, Xiner Li +211 Sept 2026q-bio.QM89
MindTopo: Can Foundation Models Reason in Topological Space?arXiv:2609.11900Yunfei Ge, Anbang Liu, Qineng Wang +211 Sept 2026cs.AI89
A Short Survey of Viewing Large Language Models in Legal AspectarXiv:2303.09136Zhongxiang Sun11 Sept 2026cs.CL89
"Mirror" Large Language Model Evaluations of Depression are Criterion ContaminatedarXiv:2508.05830Tong Li, Rasiq Hussain, Mehak Gupta +111 Sept 2026cs.CL89
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM SocietiesarXiv:2509.18052Jiaxu Zhou, Jen-tse Huang, Xuhui Zhou +211 Sept 2026cs.CL89
CHRONOBERG: Capturing Language Evolution and Temporal Awareness in Foundation ModelsarXiv:2509.22360Niharika Hegde, Subarnaduti Paul, Lars Joel-Frey +211 Sept 2026cs.CL89
Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech DetectionarXiv:2510.15685Joshua Wolfe Brook, Ilia Markov11 Sept 2026cs.CL89
Do Vision-Language Models Understand Visual Persuasiveness? A Diagnosis via Visual Persuasive FactorsarXiv:2511.17036Gyuwon Park, Hyounghun Kim11 Sept 2026cs.CL89
DeepResearch Bench II: Diagnosing Deep Research Agents via Rubrics from Expert ReportsarXiv:2601.08536Ruizhe Li, Mingxuan Du, Benfeng Xu +211 Sept 2026cs.CL89
Towards Reliable Medical LLMs: Benchmarking and Enhancing Confidence Estimation of Large Language Models in Medical ConsultationarXiv:2601.15645Zhiyao Ren, Yibing Zhan, Siyuan Liang +211 Sept 2026cs.CL89
What Language is This? Ask Your TokenizerarXiv:2602.17655Clara Meister, Ahmetcan Yavuz, Pietro Lesci +111 Sept 2026cs.CL89
Probing for Knowledge Attribution in Large Language ModelsarXiv:2602.22787Ivo Brink, Alexander Boer, Dennis Ulmer11 Sept 2026cs.CL89
Streaming Translation and Transcription Through Speech-to-Text Causal AlignmentarXiv:2603.11578Roman Koshkin, Jeon Haesung, Lianbo Liu +211 Sept 2026cs.CL89
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social InteractionarXiv:2603.17094Ryo Kamoi, Ameya Godbole, Binglin Zhou +211 Sept 2026cs.CL89
Alignment Reduces Expressed but Not Encoded Gender Bias: A Unified Framework and StudyarXiv:2603.24125Nour Bouchouchi, Thibault Laugel, Xavier Renard +211 Sept 2026cs.CL89
Analyzing LLM Reasoning to Uncover Mental Health StigmaarXiv:2604.25053Sreehari Sankar, Aliakbar Nafar, Mona Barman +211 Sept 2026cs.CL89
Timing is Everything: Temporal Scaffolding of Semantic Surprise in HumorarXiv:2605.00143Yuxi Ma, Yongqian Peng, Junchen Lyu +211 Sept 2026cs.CL89
A Recipe for Long-Context Reasoning in Large Language Models via On-Policy Optimization and DistillationarXiv:2605.12227Miguel Moura Ramos, Duarte M. Alves, Andr\'e F. T. Martins11 Sept 2026cs.CL89
Cross-lingual brain-language model alignment is robust but challenges hierarchical and computational accountsarXiv:2605.21049Ni Yang, Rui He, Philipp Homan +211 Sept 2026cs.CL89
MERIT: Matching Expertise via Rubric-Informed Training for Reviewer AssignmentarXiv:2605.27865Zixuan Yang, Yibo Zhao, Weicong Liu +111 Sept 2026cs.CL89
Characterizing Narrative Content in Web-scale LLM Pretraining DataarXiv:2606.19468Teagan Johnson, Elliott Ash, Andrew Piper +111 Sept 2026cs.CL89
Inverse Turing Bench: Evaluating Language Models as Judges of Human vs. AI DialoguearXiv:2606.21844William Hager, Ishika Rathi, Masum Hasan +111 Sept 2026cs.CL89
DiaLLM: An Investigation into the Robustness-Generation Gap in English Dialect AdaptationarXiv:2607.07669Jordan Painter, Dipankar Srirag, Adarsh Kappiyath +211 Sept 2026cs.CL89
A Factorial Study of Synthetic Data Generation for Low-Resource Machine Translation using Grammar BooksarXiv:2607.22376Varun Ghat Ravikumar, Sina Ahmadi, Lena J\"ager +111 Sept 2026cs.CL89
Predicting Startup Exit from Textual Descriptors - A Computational Linguistics FrameworkarXiv:2608.00045Alberto M. G. Saruggia, Sebastien Germano11 Sept 2026cs.CL89
Causal Episodic Memory for Feedback-Driven Agent RepairarXiv:2608.05906Khang Nhat Hoang Vo, Tam Minh Chu, Anh Trac Duc Dinh +211 Sept 2026cs.CL89
VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool UsearXiv:2608.08477Juan S. Santillana11 Sept 2026cs.CL89
Self-Evolving Embodied Agents via Skill-Harness EvolutionarXiv:2608.11350Peidong Wang, Zhiming Ma, Ying Chang +211 Sept 2026cs.CL89
Unadapted Multilingual ASR on a Garrusi Kurdish Evaluation Set: A Common-Reference Staged Normalization AnalysisarXiv:2608.16379Hiwa Asadpour11 Sept 2026cs.CL89
Aslema at NADI 2026: Data Augmentation for Intent Recognition and Slot FillingarXiv:2608.18689Tajwaar Shafiq, Hunzalah Hassan Bhatti, Firoj Alam +111 Sept 2026cs.CL89
Whitewashing Hate, Smearing Harmless Content: Annotator-Style Rebuttal Attacks on LLM-Based ModerationarXiv:2608.22230Junyu Lu, Kaiyuan Liu, Kaichun Wang +211 Sept 2026cs.CL89
DelistBench: Evaluating Search-Enabled LLMs for Auditable Corporate-Event Database CompletionarXiv:2608.22770Xuan Yao, Shuping Li, Yang Dai +211 Sept 2026cs.CL89
Toward a Cross-Lingual Romanization Ecosystem for Sinitic Languages: A Paired Mandarin-Cantonese Case StudyarXiv:2608.29170Zijie Zhang, Tan Lee, Yong Cao +111 Sept 2026cs.CL89

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Each paper page shows the abstract, related models and every source snapshot.