Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages.

510 total

Reset
Papers
TitleAuthorsPublishedCategoriesOrganization / venueQuality
EvoMaster: A Foundational Evolving Agent Framework for Agentic Science at ScalearXiv:2604.17406Bingyang Zheng, Cheng Wang, Fengyang Li +212 Sept 2026cs.AI89
Some hypotheses on how chatbots work in problem-solution-driven conversations: Large Language Models as confirmation of the Innovation IllusionarXiv:2606.07722H. D. Lethe jr, S. F. M. van Vlijmen12 Sept 2026cs.AI89
A Lightweight Multi-Agent Framework for Automated Concrete Barrier DesignarXiv:2606.12040Ran Cao, Wanting Wang, Xiye Ma +112 Sept 2026cs.AI89
Refusal Beyond a Single Direction: A Preliminary Comparison of Diff-in-Means and INLParXiv:2606.13720Alfio Ferrara, Elisabetta Rocchetti12 Sept 2026cs.AI89
Relevance Is Not Permission: Localizing and Controlling Metric-Facing Attention ContributionsarXiv:2606.30139Minwoo Yu, Young-guk Ha12 Sept 2026cs.AI89
Ceci n'est pas une pipe: AI systems as semantic abstractionsarXiv:2607.09489Jade Alglave, Patrick Cousot12 Sept 2026cs.AI89
Verification of Adaptive Agentic Controllers through Finite Rule RevisionarXiv:2607.09770Roberto Garrone12 Sept 2026cs.AI89
A Density-Matrix Framework for Electronic-Structure Analysis of Electrolytes for Lithium BatteriesarXiv:2607.25597Huize Yu, Lei Shen, Mingkang Liu12 Sept 2026cs.AI89
Improving Natural-Language Combinatorial-Optimization Accuracy in Resource-Constrained Language Models via Formal AbstractionsarXiv:2608.18409Avi Sharma, Shrenil Shaun Sharma12 Sept 2026cs.AI89
SimSkill: A Self-Evolving LLM Agent for Skill and Knowledge Accumulation in Traffic SimulationarXiv:2609.03753Can Li, Qi Liu, Qinzheng Wang +212 Sept 2026cs.AI89
HarvestBench: Measuring Whether LLM Agents Will Pay to Avoid Killing AnimalsarXiv:2609.04444Anshuman Singh, Jasmine Brazilek, Jeremiah Miller +212 Sept 2026cs.AI89
Planning and Scheduling Business Processes under Control-Flow Uncertainty: Extended VersionarXiv:2609.05578Michel Kunkler, Stefanie Rinderle-Ma12 Sept 2026cs.AI89
FastE: Readout-Triggered Token Compression for LLM Embedding InferencearXiv:2609.08407Baokun Wang, Gang Chen, Jinsong Shu +212 Sept 2026cs.AI89
GoAnt: Quality-Diversity Multi-Agent Search for Alpha Factor Discovery in Market Microstructure DataarXiv:2609.08719Stella Zhao, Tommy Sha12 Sept 2026cs.AI89
Exploring Multimodal Prompt for Visualization Authoring with Large Language ModelsarXiv:2504.13700Bo Pan, Luoxuan Weng, Minfeng Zhu +212 Sept 2026cs.HC89
A Survey of Threats Against Voice Authentication and Anti-Spoofing SystemsarXiv:2508.16843Hridoy Sankar Dutta, Kamel Kamel, Keshav Sood +112 Sept 2026cs.CR89
Spectral Masking and Interpolation Attack (SMIA): A Black-box Adversarial Attack against Voice Authentication and Anti-Spoofing SystemsarXiv:2509.07677Hridoy Sankar Dutta, Kamel Kamel, Keshav Sood +112 Sept 2026cs.SD89
Amulet: a Python Library for Assessing Interactions Among ML Defenses and RisksarXiv:2509.12386Asim Waheed, Sebastian Szyller, Vasisht Duddu12 Sept 2026cs.CR89
Generative AI performance in core undergraduate mathematics: a curriculum-level case studyarXiv:2509.13359Beatriz Navarro Lameda, Benjamin J. Walker, Nikoleta Kalaydzhieva +112 Sept 2026cs.CY89
DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive ManagementarXiv:2512.07312Chengtao Lai, Wei Zhang, Yuhang Gu +112 Sept 2026cs.AR89
Four Generations of Quantum Biomedical SensorsarXiv:2603.29944Jonathan Beaumariage, Junyu Liu, Kang Kim +212 Sept 2026quant-ph89
Causal Past Logic for Runtime Verification of Distributed LLM Agent WorkflowsarXiv:2605.20923Benedikt Bollig12 Sept 2026cs.LO89
From Agent Traces to Trust: A Survey of Evidence Tracing and Execution Provenance in LLM AgentsarXiv:2606.04990Jiaqi Zhang, Manqing Dong, Mingkai Zheng +212 Sept 2026cs.CR89
Adaptive Perturbation Selection for Contrastive Audio DecodingarXiv:2607.00247Aaron Isidore Grace, Weiran Wang, Zhouyuan Huo12 Sept 2026cs.SD89
Natural Language Access to Domain-Specific Metadata: A Reusable Framework for LLM Query GenerationarXiv:2607.18029Blake G. Fitch, Cato Elia Kurtz12 Sept 2026cs.DB89
GitSkills: A Dataset of Agent Skills on GitHubarXiv:2608.10906Daniel Graziotin, Giuseppe Destefanis, Marco Ortu +112 Sept 2026cs.SE89
FaultLens: Learning Compact Behavioral Test Suites for Generated Operational ProgramsarXiv:2608.26746Hang Lyu, Jingtao Zhang, Zeming Liu12 Sept 2026cs.SE89
OmegaUse-SOP: SOP Engineering for Professional Computer Use from Human DemonstrationsarXiv:2609.02149Hua Wu, Hucheng Yang, Jingbo Zhou +212 Sept 2026cs.HC89
Toward Collective-Centric Evaluation of Preference Inference for Participatory DemocracyarXiv:2609.02990Benjamin Piwowarski, David Mas, Fran\c{c}ois Yvon +212 Sept 2026cs.SI89
Ordinary, Reasonable Chatbots: Do AI Models Track Human Legal Judgments?arXiv:2609.06769Christopher Buccafusco, Emily Wenger, Nirav Patel12 Sept 2026cs.CY89
Monadic Second-Order Logic in HOL: Deep and Shallow with Automated Faithfulness (Extended Preprint)arXiv:2609.07345Christoph Benzmueller, Daniel Kirchner12 Sept 2026cs.LO89
Your Agent Says Yes: Interpreting Adversarial Market Behavior Beyond Individual TransactionsarXiv:2609.07675Matt White, Tianyu Shi, Xiao-Yang Liu +212 Sept 2026cs.CE89
A Multi-Modal Perception Pipeline for Object Detection and Tracking in Autonomous RacingarXiv:2609.08338Ayoub Raji, Davide Malvezzi, Fabio Bagni +212 Sept 2026cs.RO89
OpenDiscoveryTrace: Process Traces for Evaluating AI Scientist WorkflowsarXiv:2609.09203Aayam Bansal, Keertan Balaji11 Sept 2026cs.AI89
Adaptive Entangled Game Modules in Artificial General IntelligencearXiv:2609.09226Haochen Li, Xinshuai Guo, Jingdong Ouyang +211 Sept 2026cs.AI89
Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic TasksarXiv:2609.09233Wasu Top Piriyakulkij, Rachel Lawrence, Alicia Curth +211 Sept 2026cs.AI89
Gradland: On Phenomenal Experience, Differentiated Across Many DimensionsarXiv:2609.09306David Balduzzi11 Sept 2026cs.AI89
An Autonomous GeoAI Agent for Arctic Eco-NavigationarXiv:2609.09374Samira Alkaee Taleghan, Younghyun Koo, Farnoush Banaei-Kashani11 Sept 2026cs.AI89
The Menu Is an Execution Prior: State-Path Tool Menus for Online AgentsarXiv:2609.09395Bo Yan, Weikai Lin, Song Wang11 Sept 2026cs.AI89
Decision-Focused Active Learning for Scale-Aware Critical-Materials RecoveryarXiv:2609.09413Niranjan Srinivas, Debajyoti Ray, Elias Nakouzi11 Sept 2026cs.AI89

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Each paper page shows the abstract, related models and every source snapshot.