Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

1,189 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
Which Medical Questions Deserve Rationales? Perturbation-Sensitive Selection for Robust QAarXiv:2609.09684cs.CLYuexin Wu, Dayou Yu, Vasile Rus11 Sept 2026
Looped GPT-BERT: Trading Parameters for Computation in Small Language ModelingarXiv:2609.09691cs.CLTingshuo Fan, Hongtao Mu, Tianyu Zhou +211 Sept 2026
CT-SAFR: Safe and Interpretable Chain-of-Thought Reasoning for Autonomous Robots: A Multi-Layered Verification Framework for Trustworthy AI-Driven Robotic Decision MakingarXiv:2609.09692cs.ROCagri Temel11 Sept 2026
When Auditors Fabricate: Batch-Size Degradation and Confident Hallucination in LLM Detection of Planted Document ContaminationarXiv:2609.09696cs.CLKaran Parekh, Sanjana Pendyala Ravinder, Sana Mhapsekar +111 Sept 2026
Kernel-Complexity Edge Sanitization for Training-Free Defense against Structural Graph AttacksarXiv:2609.09698cs.LGYaning Jia, Shenyang Deng, Yaoqing Yang +211 Sept 2026
Distilling Image Prototypes for Guided Test-Time AdaptationarXiv:2609.09737cs.CVLiwen Wang, Xingbo Dong, Iman Yi Liao +211 Sept 2026
HiRAD: A Flexible Large-Scale AGV Routing SystemarXiv:2609.09752cs.ROYunjie Huang, Ruizhong Wu, Mengxuan Zhang +211 Sept 2026
Fine-Tuning a KV Cache Concatenation-Aware Model or Recomputing KV Caches? Why Not Both?arXiv:2609.09768cs.LGFumihiko Tachibana, Daisuke Miyashita, Jun Deguchi11 Sept 2026
BRACE: Anchored Bellman-Residual Correction for Stale Critics in Asynchronous RLarXiv:2609.09783cs.LGGuanqun Zhao, Zijun Xie, Binbin Zheng +211 Sept 2026
Pairit: A Platform for Live Experiments on Human-AI CollaborationarXiv:2609.09789cs.HCHarang Ju, Sinan Aral11 Sept 2026
LogiScope-VQA: Benchmarking Vision-Language Models for Logistics Hazard Identification in Industrial ScenariosarXiv:2609.09790cs.CVHanjing Zhou, Mingze Yin, Ying Lian +211 Sept 2026
How Fragile Is Safety Alignment at Frontier Scale? A Single-Direction Attack on a 320B MoEarXiv:2609.09793cs.CRYi Shi, Tanyu Chen, Kai Shen11 Sept 2026
CS-Guard: Benchmarking LLM Guardrails for Code Generation SecurityarXiv:2609.09798cs.CRJinyang Li, Mingyu Guo, Hung X. Nguyen11 Sept 2026
uFlowCSP: Crystal Structure Prediction using Mean flow generative modelsarXiv:2609.09799cond-mat.mtrl-sciSourin Dey, Dipannoy Das Gupta, Lai Wei +211 Sept 2026
Subgroup Membership Inference Audits of Differentially Private Synthetic TextarXiv:2609.09848cs.CRYidan Sun, Viktor Schlegel, Srinivasan Nandakumar +211 Sept 2026
Can AI Agents Detect and Repair Artifact Drift in Network Experiments?arXiv:2609.09849cs.NITianzhu Zhang, Weichen Tao, Changgang Zheng +211 Sept 2026
With a Thermomix You Lose the Ability to Cook: A Kitchen Machine Analogy for Applications of Generative AI in EducationarXiv:2609.09856cs.CYNikol Rummel, Valentina Nachtigall, Ernesto Panadero11 Sept 2026
Forward-Free LLM Depth Pruning via Weight RedundancyarXiv:2609.09883cs.LGVincent-Daniel Yun, Woosang Lim11 Sept 2026
Albedo Estimation via Latent Bridge MatchingarXiv:2609.09884cs.CVCarme Corbi, David Serrano-Lozano, Javier Vazquez-Corral +111 Sept 2026
Strangers to Themselves: What Language Models Say About Themselves Is GenericarXiv:2609.09899cs.LGPhil Blandfort, Urja Pawar11 Sept 2026
FlowCPO: A Unified Divergence View of Preference Alignment for Flow ModelsarXiv:2609.09905stat.MLYansen Han, Shengyi Liao, Peng Sun +211 Sept 2026
Improving Cross-Lingual Token Representations by Adding a Pinch of SALTarXiv:2609.09953cs.CLGuillem Ram\'irez11 Sept 2026
Fidelity-Aware Scheduling of Quantum Circuits on Multi-QPU SystemsarXiv:2609.09980quant-phInnocenzo Fulginiti, Antonio Tudisco, Salvatore Zammuto +211 Sept 2026
What Makes Adversarial Examples Transfer Across Deepfake Detectors?arXiv:2609.10002cs.CVRafael M. Mamede, Pedro C. Neto, Ana F. Sequeira11 Sept 2026
MetroLLM-Bench: Evaluating Language Models as Transit Kiosk RuntimesarXiv:2609.10016cs.LGRemco Hendriks (Continker)11 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.