Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

1,189 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
IBIB: A Protocol for Measuring Enterprise AI Systems by Serving Route, Not Model IdentifierarXiv:2609.10494cs.CLBlake Stenstrom, Charangan Vasantharajan, Brian Sathianathan11 Sept 2026
Show-Harness: Just a VLM Agent Can Play RobotsarXiv:2609.10522cs.ROYanzhe Chen, Zechen Bai, Zhijun Cao +211 Sept 2026
Reinforcement Learning with Temporal-Logic-Based Causal DiagramsarXiv:2306.13732cs.AIYash Paliwal, Rajarshi Roy, Jean-Rapha\"el Gaglione +211 Sept 2026
Reinforcement learning for Quantum Tiq-Taq-ToearXiv:2411.06429cs.AICatalin-Viorel Dinu, Thomas Moerland11 Sept 2026
ROTATE: Regret-driven Open-ended Training for Ad Hoc TeamworkarXiv:2505.23686cs.AICaroline Wang, Arrasy Rahman, Benjamin Nativi +211 Sept 2026
RelayS2S: A Dual-Path Speculative Generation for Real-Time DialoguearXiv:2603.23346cs.AILong Mai, Junli Liang11 Sept 2026
MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory PredictionarXiv:2604.10169cs.AIWenchang Duan, Zhenguo Gao, Jinguo Xian +111 Sept 2026
Zero-shot World Models Are Developmentally Efficient LearnersarXiv:2604.10333cs.AIKhai Loong Aw, Klemen Kotar, Wanhee Lee +211 Sept 2026
Non-Stationarity Breaks Permutation Surrogates in Multi-Agent Reinforcement Learning: Diagnosis and RemediesarXiv:2604.23716cs.AINikolaos Al. Papadopoulos, Konstantinos E. Psannis11 Sept 2026
CoGReV: A Confidence-Gated Post-Hoc Non-Monotonic Belief Revision Framework for Phishing Website ClassificationarXiv:2604.25512cs.AIMainak Sen, Kumar Sankar Ray, Amlan Chakrabarti11 Sept 2026
Grounded Continuation: A Linear-Time Runtime Verifier for LLM ConversationsarXiv:2605.14175cs.AIQisong He, Jinwei Hu, Xinmiao Huang +211 Sept 2026
Cultural Binding Heads in Language ModelsarXiv:2605.28543cs.AIAvrile Floro, Luca Benedetto11 Sept 2026
KairosAgent: Agentic Time Series Forecasting with Fused Semantic ReasoningarXiv:2605.30002cs.AIKun Feng, Ziwei Shan, Yuchen Fang +211 Sept 2026
Self-Evolving Scientific Agent Designs Physically Reasoned White-Box Fluid ControlarXiv:2606.08405cs.AIBoai Sun, Wenjin Guo, Zongmin Yu +111 Sept 2026
KernelGenBench: Can LLMs and Agents Write Efficient Kernels Across Operator Sources and Hardware Platforms?arXiv:2607.27231cs.AIPeiyu Zang, Jian Tao, Jialing Zhang +211 Sept 2026
ViSR-KGC: Visual Subgraph Reasoning with Vision-Language Models for Multimodal Knowledge Graph CompletionarXiv:2608.05833cs.AIJiafan Li, Mengxue Yang, Jiaqi Zhu +211 Sept 2026
Learning to Predict Middle-Layer Attention in MLLMs for Visual Token PruningarXiv:2608.06411cs.AIYuyao Sun, Tao Deng, Shuang Li +211 Sept 2026
LiFTER: A Grounded Neuro-Symbolic Microscope for Continuous-Time Dynamic Graph ForecastingarXiv:2608.06765cs.AIMinwoo Yu, Young-guk Ha11 Sept 2026
A Human Audit of OpenAIs AI-Generated Mathematical ProofsarXiv:2608.14673cs.AIMiko{\l}aj Sienicki, Krzysztof Sienicki11 Sept 2026
Dear Algo: A Precision-First Agentic Intent Layer for Unified Search and RecommendationarXiv:2608.15877cs.AIRui Wang, Jiazhou Wang, Zheng Wei +211 Sept 2026
Physics of Agents: Statistical Mechanics Predicts Collective Behavior of AI AgentsarXiv:2608.16578cs.AIBatu El, Jinhee Paeng, Fatih Dinc +211 Sept 2026
FrontierChallenge: Evaluating Scientific Workflow CompletionarXiv:2608.24979cs.AILiangcai Su, Zhaopeng Feng, Zhuo Chen +211 Sept 2026
A Composable Evaluation System for Reproducible Omni-Modal Foundation Model EvaluationarXiv:2609.01315cs.AIHodong Lee, Sanghee Park, Dohoon Ryu +211 Sept 2026
Beyond Prompts: Measuring and Optimizing LLM Tool-Agent HarnessesarXiv:2609.05736cs.AICen Mia Zhao, Haibo Ruan, Wenjie Chen +211 Sept 2026
From Monolithic Blending to Agentic Orchestration: Dynamic Response for Conversational Assistants at ScalearXiv:2609.05758cs.AICen Mia Zhao, Peng Wang, Chuan Shi +211 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.