Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

210 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
Analyzing Traditional and Neural Approaches to Multilingual Readability AssessmentarXiv:2609.10792cs.CLJoshua Wong, Chris Tanner11 Sept 2026
Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error CorrectionarXiv:2609.10810cs.CLKateryna Karpo, Artem Chernodub11 Sept 2026
LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease DetectionarXiv:2609.10896cs.CLXiao Wei, Yuqin Lin, Yaru Cao +211 Sept 2026
SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query GraphsarXiv:2609.10901cs.CLJiacheng Sang, Mengyuan Li, Sanxing Chen +211 Sept 2026
Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender SystemarXiv:2609.10922cs.CLMing Li, Dai Li, Xuying Ning +211 Sept 2026
Using Semantic Uncertainty to Estimate Transition Relevance in Turn-takingarXiv:2609.10934cs.CLMuhammad Umair, Jan P. de Ruiter11 Sept 2026
Distribution-aware Language Neuron Identification in Multilingual Large Language ModelsarXiv:2609.10993cs.CLMinjun Kim, Inho Won, Junghun Yuk +211 Sept 2026
Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary ModelsarXiv:2609.10996cs.CLYu-Chung Hsiao11 Sept 2026
K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language ModelsarXiv:2609.11020cs.CLYu Sun, Mengyin Lu, Cong Feng +211 Sept 2026
ProMediConv: Benchmarking Proactive Conversational Agents in Legal Dispute MediationarXiv:2609.11101cs.CLZesheng Wei, Mengfan Li, Wenhao Liu +211 Sept 2026
Overview of the NLPCC 2026 Shared Task 11: Agent-Based Experiment Reproduction from Scientific PapersarXiv:2609.11117cs.CLHanhua Hong, Yizhi Li, Luu Gia Huy +211 Sept 2026
From Repetition to Recognition: Inductive Discovery of Disinformation NarrativesarXiv:2609.11128cs.CLMax Upravitelev, Veronika Solopova, Jing Yang +211 Sept 2026
Rubric-Aligned Disentangled Evaluation of Human Simultaneous InterpretingarXiv:2609.11131cs.CLZiyu Zhang, Satoshi Nakamura11 Sept 2026
Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema NormalizationarXiv:2609.11141cs.CLDong-Jae Koh, Huisu Kim, SeongHwan Yoon +211 Sept 2026
FlexComp: One Model for Every Ratio in Context CompressionarXiv:2609.11192cs.CLKaiyan Zhao, Zhongtao Miao, Akiko Aizawa +111 Sept 2026
Automated Identification of Competing Narratives in Political Discourse on Social MediaarXiv:2609.11202cs.CLSergej Wildemann, Erick Elejalde11 Sept 2026
OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language ModelsarXiv:2609.11244cs.CLJianjiang Yang, Peihang Li, Shanqing Xu +211 Sept 2026
Assessing the Reusability of Public Speech Resources for Low-Resource Languages: A Central Kurdish Case StudyarXiv:2609.11246cs.CLHiwa Asadpour11 Sept 2026
The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based MethodsarXiv:2609.11247cs.CLIoanna Kaffeza, Efthymios Georgiou, Alexandros Potamianos11 Sept 2026
Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper AdaptationarXiv:2609.11302cs.CLMaria Frangiadaki, Dimitrios Damianos, Kosmas Kritsis +111 Sept 2026
MultiHuSE: A Multimodal Dataset for Humour Styles and EmotionsarXiv:2609.11322cs.CLMary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat11 Sept 2026
SEAR: Segment-Evidence-Aware Routing for Weak-to-Strong Multilingual Speech MCQarXiv:2609.11355cs.CLHuy Hoang Le, Long-Bao Nguyen, Minh Tri Dao11 Sept 2026
TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model OutputsarXiv:2609.11399cs.CLShenbin Qian, Yves Scherrer11 Sept 2026
ReGround: Grounding Reviewer Comments in Multimodal EvidencearXiv:2609.11460cs.CLSerwar Basch, Lizhen Qu, Iryna Gurevych11 Sept 2026
Complex-Text Robustness Evaluation and Failure Diagnosis for Low-Resource Multilingual Text-to-SpeecharXiv:2609.11545cs.CLTianlun Zuo, Ziyu Zhang, Tingzhi Mao +211 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.