Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

1,189 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity LinkingarXiv:2609.10745cs.CLParinthapat Pengpun, Simran Khanuja, Graham Neubig11 Sept 2026
Analyzing Traditional and Neural Approaches to Multilingual Readability AssessmentarXiv:2609.10792cs.CLJoshua Wong, Chris Tanner11 Sept 2026
Larger Context Window, Fewer Overcorrections: Optimizing Prompts and Batching for Minimal-Edit Grammatical Error CorrectionarXiv:2609.10810cs.CLKateryna Karpo, Artem Chernodub11 Sept 2026
LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease DetectionarXiv:2609.10896cs.CLXiao Wei, Yuqin Lin, Yaru Cao +211 Sept 2026
SearchAtlas: Analyzing Agentic Search Strategies via Evidential Query GraphsarXiv:2609.10901cs.CLJiacheng Sang, Mengyuan Li, Sanxing Chen +211 Sept 2026
Auto-RecSys: Harnessing Autonomous Research Agents for Industry-Scale Recommender SystemarXiv:2609.10922cs.CLMing Li, Dai Li, Xuying Ning +211 Sept 2026
Using Semantic Uncertainty to Estimate Transition Relevance in Turn-takingarXiv:2609.10934cs.CLMuhammad Umair, Jan P. de Ruiter11 Sept 2026
Distribution-aware Language Neuron Identification in Multilingual Large Language ModelsarXiv:2609.10993cs.CLMinjun Kim, Inho Won, Junghun Yuk +211 Sept 2026
Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary ModelsarXiv:2609.10996cs.CLYu-Chung Hsiao11 Sept 2026
K/V-Cache Interventions Dissociate Representation Alignment from Persona Expression in Decoder-Only Language ModelsarXiv:2609.11020cs.CLYu Sun, Mengyin Lu, Cong Feng +211 Sept 2026
ProMediConv: Benchmarking Proactive Conversational Agents in Legal Dispute MediationarXiv:2609.11101cs.CLZesheng Wei, Mengfan Li, Wenhao Liu +211 Sept 2026
Overview of the NLPCC 2026 Shared Task 11: Agent-Based Experiment Reproduction from Scientific PapersarXiv:2609.11117cs.CLHanhua Hong, Yizhi Li, Luu Gia Huy +211 Sept 2026
From Repetition to Recognition: Inductive Discovery of Disinformation NarrativesarXiv:2609.11128cs.CLMax Upravitelev, Veronika Solopova, Jing Yang +211 Sept 2026
Rubric-Aligned Disentangled Evaluation of Human Simultaneous InterpretingarXiv:2609.11131cs.CLZiyu Zhang, Satoshi Nakamura11 Sept 2026
Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema NormalizationarXiv:2609.11141cs.CLDong-Jae Koh, Huisu Kim, SeongHwan Yoon +211 Sept 2026
FlexComp: One Model for Every Ratio in Context CompressionarXiv:2609.11192cs.CLKaiyan Zhao, Zhongtao Miao, Akiko Aizawa +111 Sept 2026
Automated Identification of Competing Narratives in Political Discourse on Social MediaarXiv:2609.11202cs.CLSergej Wildemann, Erick Elejalde11 Sept 2026
OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language ModelsarXiv:2609.11244cs.CLJianjiang Yang, Peihang Li, Shanqing Xu +211 Sept 2026
Assessing the Reusability of Public Speech Resources for Low-Resource Languages: A Central Kurdish Case StudyarXiv:2609.11246cs.CLHiwa Asadpour11 Sept 2026
The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based MethodsarXiv:2609.11247cs.CLIoanna Kaffeza, Efthymios Georgiou, Alexandros Potamianos11 Sept 2026
Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper AdaptationarXiv:2609.11302cs.CLMaria Frangiadaki, Dimitrios Damianos, Kosmas Kritsis +111 Sept 2026
MultiHuSE: A Multimodal Dataset for Humour Styles and EmotionsarXiv:2609.11322cs.CLMary Ogbuka Kenneth, Foaad Khosmood, Abbas Edalat11 Sept 2026
SEAR: Segment-Evidence-Aware Routing for Weak-to-Strong Multilingual Speech MCQarXiv:2609.11355cs.CLHuy Hoang Le, Long-Bao Nguyen, Minh Tri Dao11 Sept 2026
TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model OutputsarXiv:2609.11399cs.CLShenbin Qian, Yves Scherrer11 Sept 2026
ReGround: Grounding Reviewer Comments in Multimodal EvidencearXiv:2609.11460cs.CLSerwar Basch, Lizhen Qu, Iryna Gurevych11 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.