Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

210 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
A Training-Free, Alignment-Free Approach to Corporate Intelligence: Application to SEC FilingsarXiv:2609.11620cs.CLJean-Fran\c{c}ois Delpech11 Sept 2026
Structured Transforms for Low-Overhead Quantization of Language ModelsarXiv:2609.11687cs.CLDaria Cherniuk, Alexander Rudikov, Boris Kashin +111 Sept 2026
The Eloquence submission for Task 2 of the Interspeech 2026 MLC-SLM challengearXiv:2609.11724cs.CLJordi Luque, Lorenzo Concina, Marco Matassoni +211 Sept 2026
RAG-Safety-Bench: Reliable Evaluation of Retrieval-Augmented LLM SafetyarXiv:2609.11758cs.CLAdithiyan Rajan Indira Saravanan, Kathleen C. Fraser11 Sept 2026
Component-Aware Differential Privacy for Federated Multilingual Speech-LLMsarXiv:2609.11762cs.CLJordi Luque, Fernando L\'opez, Aleix Sant11 Sept 2026
The widening evaluation gap in medical large language model research 2023 to 2026arXiv:2609.11770cs.CLRaad Bin Tareaf, Murad Al-Rajab, Samia Loucif11 Sept 2026
Target leakage, not model class, explains reported accuracy in survey-based cardiovascular screening: a leakage-tiered audit of glass-box and tabular foundation modelsarXiv:2609.11838cs.CLRaad Bin Tareaf, Murad Al-Rajab, Samia Loucif +211 Sept 2026
IndicTriMix: Developing Language Identification Datasets and Models for Tri-Language Code-MixingarXiv:2609.11851cs.CLPruthwik Mishra, Rudra Trivedi, Avi Patel +211 Sept 2026
Epistemic orientation predicts legislative effectiveness among members of the US CongressarXiv:2609.11865cs.CLSegun Aroyehun, Stephan Lewandowsky, David Garcia11 Sept 2026
Augustinian BabyLM: What Ostensive Definition Can and Cannot Teach a Small Language ModelarXiv:2609.11870cs.CLLisa Bylinina11 Sept 2026
Nuha-Speech: Building General-Purpose Arabic Speech-LLMsarXiv:2609.11892cs.CLYingzhi Wang, Reem Alhazzani, Muhammad Alqurishi11 Sept 2026
Distance generalization in transformers: why bother with positional encoding?arXiv:2609.11913cs.CLDaniel Henrik Nevermann, Claudius Gros11 Sept 2026
More than half of recent astronomy papers are written with language-model assistancearXiv:2609.10664astro-ph.IMSerat M. Saad, Yuan-Sen Ting11 Sept 2026
BodyCam-VQA: Enhanced Body-Worn Camera Video Captioning via Multimodal Reasoning and Probe Question GenerationarXiv:2609.10815cs.CVKarish Gupta, Matthew Alex, Alex Li +211 Sept 2026
INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry ArchivesarXiv:2609.11261cs.DLDaniel Akselrad, Robert N. Proctor11 Sept 2026
Xiaomi-CocktailASR-1 Technical ReportarXiv:2609.11274cs.SDYiru Zhang, Hang Su, Lichun Fan +211 Sept 2026
Whisper-Based Speech Transcription from Videos Across Multiple Languages for Cross-Cultural UnderstandingarXiv:2609.11772eess.ASMichael Picheny11 Sept 2026
SpecGuard: Inference-Time Backdoor Detection For FreearXiv:2609.11799cs.CRRui Wen, Ahmed Salem, Andrew Paverd +211 Sept 2026
A Short Survey of Viewing Large Language Models in Legal AspectarXiv:2303.09136cs.CLZhongxiang Sun11 Sept 2026
"Mirror" Large Language Model Evaluations of Depression are Criterion ContaminatedarXiv:2508.05830cs.CLTong Li, Rasiq Hussain, Mehak Gupta +111 Sept 2026
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM SocietiesarXiv:2509.18052cs.CLJiaxu Zhou, Jen-tse Huang, Xuhui Zhou +211 Sept 2026
Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech DetectionarXiv:2510.15685cs.CLJoshua Wolfe Brook, Ilia Markov11 Sept 2026
Do Vision-Language Models Understand Visual Persuasiveness? A Diagnosis via Visual Persuasive FactorsarXiv:2511.17036cs.CLGyuwon Park, Hyounghun Kim11 Sept 2026
DeepResearch Bench II: Diagnosing Deep Research Agents via Rubrics from Expert ReportsarXiv:2601.08536cs.CLRuizhe Li, Mingxuan Du, Benfeng Xu +211 Sept 2026
Towards Reliable Medical LLMs: Benchmarking and Enhancing Confidence Estimation of Large Language Models in Medical ConsultationarXiv:2601.15645cs.CLZhiyao Ren, Yibing Zhan, Siyuan Liang +211 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.