Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages.

210 total

Reset
Papers
TitleAuthorsPublishedCategoriesOrganization / venueQuality
Reason Through the Latent! Making Latent Visual Reasoning NecessaryarXiv:2609.06746Suhyeong Park, Junha Jung, Jaewoo Kang11 Sept 2026cs.AI89
Limitations of Automated Simulatability: LLM Simulators Can Bypass ExplanationsarXiv:2609.08585Antonin Poch\'e, Fanny Jourdan, Nils Feldhus +211 Sept 2026cs.CL89
Data-Efficient Language Modeling: From Frontier Advancement to Principle-Guided Model ImprovementarXiv:2609.10702Shuxing Yang, Kaihao Zhu, Junjie Yang +211 Sept 2026cs.CL89
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept PredictionarXiv:2609.10715NCP Team, Jiaqi Cao, Chiyu Chen +211 Sept 2026cs.CL84
CMNIE: An Information Extraction Benchmark for Chinese Military NewsarXiv:2609.10722Yan Yu, Mengna Zhu, Zhenyu Song +211 Sept 2026cs.CL89
Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity LinkingarXiv:2609.10745Parinthapat Pengpun, Simran Khanuja, Graham Neubig11 Sept 2026cs.CL81
Analyzing Traditional and Neural Approaches to Multilingual Readability AssessmentarXiv:2609.10792Joshua Wong, Chris Tanner11 Sept 2026cs.CL89
Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error CorrectionarXiv:2609.10810Kateryna Karpo, Artem Chernodub11 Sept 2026cs.CL89
Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current ArchitecturesarXiv:2609.10893Victor Mazzotti, Luiz Pereira, Marina Bitencourt dos Santos +211 Sept 2026cs.CL89
LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease DetectionarXiv:2609.10896Xiao Wei, Yuqin Lin, Yaru Cao +211 Sept 2026cs.CL89
SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query GraphsarXiv:2609.10901Jiacheng Sang, Mengyuan Li, Sanxing Chen +211 Sept 2026cs.CL89
Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender SystemarXiv:2609.10922Ming Li, Dai Li, Xuying Ning +211 Sept 2026cs.CL89
Using Semantic Uncertainty to Estimate Transition Relevance in Turn-takingarXiv:2609.10934Muhammad Umair, Jan P. de Ruiter11 Sept 2026cs.CL89
Distribution-aware Language Neuron Identification in Multilingual Large Language ModelsarXiv:2609.10993Minjun Kim, Inho Won, Junghun Yuk +211 Sept 2026cs.CL89
Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary ModelsarXiv:2609.10996Yu-Chung Hsiao11 Sept 2026cs.CL89
K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language ModelsarXiv:2609.11020Yu Sun, Mengyin Lu, Cong Feng +211 Sept 2026cs.CL89
ProMediConv: Benchmarking Proactive Conversational Agents in Legal Dispute MediationarXiv:2609.11101Zesheng Wei, Mengfan Li, Wenhao Liu +211 Sept 2026cs.CL89
Overview of the NLPCC 2026 Shared Task 11: Agent-Based Experiment Reproduction from Scientific PapersarXiv:2609.11117Hanhua Hong, Yizhi Li, Luu Gia Huy +211 Sept 2026cs.CL89
From Repetition to Recognition: Inductive Discovery of Disinformation NarrativesarXiv:2609.11128Max Upravitelev, Veronika Solopova, Jing Yang +211 Sept 2026cs.CL89
Rubric-Aligned Disentangled Evaluation of Human Simultaneous InterpretingarXiv:2609.11131Ziyu Zhang, Satoshi Nakamura11 Sept 2026cs.CL89
Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema NormalizationarXiv:2609.11141Dong-Jae Koh, Huisu Kim, SeongHwan Yoon +211 Sept 2026cs.CL89
FlexComp: One Model for Every Ratio in Context CompressionarXiv:2609.11192Kaiyan Zhao, Zhongtao Miao, Akiko Aizawa +111 Sept 2026cs.CL89
Automated Identification of Competing Narratives in Political Discourse on Social MediaarXiv:2609.11202Sergej Wildemann, Erick Elejalde11 Sept 2026cs.CL89
OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language ModelsarXiv:2609.11244Jianjiang Yang, Peihang Li, Shanqing Xu +211 Sept 2026cs.CL89
Assessing the Reusability of Public Speech Resources for Low-Resource Languages: A Central Kurdish Case StudyarXiv:2609.11246Hiwa Asadpour11 Sept 2026cs.CL89
The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based MethodsarXiv:2609.11247Ioanna Kaffeza, Efthymios Georgiou, Alexandros Potamianos11 Sept 2026cs.CL89
Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper AdaptationarXiv:2609.11302Maria Frangiadaki, Dimitrios Damianos, Kosmas Kritsis +111 Sept 2026cs.CL89
MultiHuSE: A Multimodal Dataset for Humour Styles and EmotionsarXiv:2609.11322Mary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat11 Sept 2026cs.CL89
On the Impact of Anonymization on the Performance of Large Language ModelsarXiv:2609.11335Tobias Deu{\ss}er, Max Hahnb\"uck, Lorenz Sparrenberg +211 Sept 2026cs.CL89
SEAR: Segment-Evidence-Aware Routing for Weak-to-Strong Multilingual Speech MCQarXiv:2609.11355Huy Hoang Le, Long-Bao Nguyen, Minh Tri Dao11 Sept 2026cs.CL89
TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model OutputsarXiv:2609.11399Shenbin Qian, Yves Scherrer11 Sept 2026cs.CL89
SWRouter: Similarity-Contractive Window Routing for Multi-Turn Large Language Model ConversationsarXiv:2609.11414Yu Wang, Yuchen Li, Rui Kong +211 Sept 2026cs.CL89
Cross-Lingual Clinical Annotation Projection as Constrained Text Generation: A Six-Language StudyarXiv:2609.11450\'Alvaro Rey-Blanes, Francisco J. Moreno-Barea, Francisco J. Veredas11 Sept 2026cs.CL89
ReGround: Grounding Reviewer Comments in Multimodal EvidencearXiv:2609.11460Serwar Basch, Lizhen Qu, Iryna Gurevych11 Sept 2026cs.CL89
Complex-Text Robustness Evaluation and Failure Diagnosis for Low-Resource Multilingual Text-to-SpeecharXiv:2609.11545Tianlun Zuo, Ziyu Zhang, Tingzhi Mao +211 Sept 2026cs.CL89
A Training-Free, Alignment-Free Approach to Corporate Intelligence: Application to SEC FilingsarXiv:2609.11620Jean-Fran\c{c}ois Delpech11 Sept 2026cs.CL89
Structured Transforms for Low-Overhead Quantization of Language ModelsarXiv:2609.11687Daria Cherniuk, Alexander Rudikov, Boris Kashin +111 Sept 2026cs.CL89
The Eloquence submission for Task 2 of the Interspeech 2026 MLC-SLM challengearXiv:2609.11724Jordi Luque, Lorenzo Concina, Marco Matassoni +211 Sept 2026cs.CL89
RAG-Safety-Bench: Reliable Evaluation of Retrieval-Augmented LLM SafetyarXiv:2609.11758Adithiyan Rajan Indira Saravanan, Kathleen C. Fraser11 Sept 2026cs.CL89
Component-Aware Differential Privacy for Federated Multilingual Speech-LLMsarXiv:2609.11762Jordi Luque, Fernando L\'opez, Aleix Sant11 Sept 2026cs.CL89

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Each paper page shows the abstract, related models and every source snapshot.