Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

1,189 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
False positive bias in AI-powered speech-based cognitive screening for multilingual English speakers in the UKarXiv:2602.13047cs.CLMadhurananda Pahar, Caitlin Illingworth, Dorota Braun +211 Sept 2026
City Editing: Hierarchical Agentic Execution for Dependency-Aware Urban Geospatial ModificationarXiv:2602.19326cs.MARui Liu, Steven Jige Quan, Zhong-Ren Peng +211 Sept 2026
Spec-Harness: Measuring and Improving Behavioral Adequacy of LLM-Synthesized Formal SpecificationsarXiv:2604.00280cs.SEMd Rakib Hossain Misu, Iris Ma, Cristina V. Lopes11 Sept 2026
Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement LearningarXiv:2604.10701cs.LGZikang Shan, Han Zhong, Liwei Wang +111 Sept 2026
Where is the Mind? Persona Vectors and LLM IndividuationarXiv:2604.17031cs.CLPierre Beckmann, Patrick Butlin11 Sept 2026
The Biggest Risk of Embodied AI is Governance LagarXiv:2604.21938cs.CYShaoshan Liu11 Sept 2026
Dont Just Teach, Explain! A Gamified 20Q Recommender for Cybersecurity EducationarXiv:2604.26964cs.CYMary Nusrat, Sarfuddin Bhuiyan, Gahangir Hossain11 Sept 2026
"What Are You Really Trying to Do?": Co-Creating Life Goals from Everyday Computer UsearXiv:2605.00497cs.HCShardul Sapkota, Matthew J\"orke, Zane Sabbagh +211 Sept 2026
EVA-Bench: A New End-to-end Framework for Evaluating Voice AgentsarXiv:2605.13841cs.SDTara Bogavelli, Gabrielle Gauthier Melan\c{c}on, Katrina Stankiewicz +211 Sept 2026
Complementing reinforcement learning with SFT through logit averaging in the post training of LLMsarXiv:2605.20555cs.LGXingwei Gan, Ying Zhu11 Sept 2026
SpecBench: Measuring Reward Hacking in Long-Horizon Coding AgentsarXiv:2605.21384cs.SEBingchen Zhao, Dhruv Srikanth, Yuxiang Wu +111 Sept 2026
Tracing Computation Density in LLMsarXiv:2605.27033cs.CLCorentin Kervadec, Iuliia Lysova, Iuri Macocco +211 Sept 2026
BaltiVoice: A Speech Corpus and Fine-tuned Whisper ASR System for the Balti LanguagearXiv:2606.03504cs.CLMuhammad Ali11 Sept 2026
Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement LearningarXiv:2606.03962cs.LGAnthony GX-Chen, Ankit Anand, Gheorghe Comanici +211 Sept 2026
FP8 is All You Need (Part 1): Debunking Hardware FP64 as the HPC Holy Grail (Sep 3rd version)arXiv:2606.06510cs.ARSatoshi Matsuoka11 Sept 2026
FiberTune: Preserving Action-Fiber Visual Residuals in Vision-Language-Action Fine-TuningarXiv:2606.08653cs.CVHaihao Lin, Xiangsheng Huang, Xiao Yang +211 Sept 2026
Expert-Level Crisis Detection in Mental Health ConversationsarXiv:2606.10380cs.CLGrace Byun, Abigail Lott, Rebecca Lipschutz +211 Sept 2026
PSCT-Net: Geometry-Aware Pediatric Skull CT Reconstruction via Differentiable Back-Projection and Attention-Guided RefinementarXiv:2606.19867cs.CVDong Yeong Kim, Jaewon Choi, Youmin Shin +211 Sept 2026
FP8 is All You Need (Part 2): Full-FP64 3-D FFT on FP8-Generation Tensor CoresThe Integer-Epilogue Wall and the Minimal Hardware That Would Remove ItarXiv:2606.23698cs.MSSatoshi Matsuoka11 Sept 2026
Spectral Geometry and Bosonic-Bloch Probes: Explorations in Quantum LearningarXiv:2607.00063quant-phSantanu Ganguly, Xing Liang, Dimitrios Makris11 Sept 2026
Builder, Defender, Breaker: Measurable Independence and Bounded Autonomy When Generative Models Build, Defend and Test SoftwarearXiv:2607.03215cs.CRMohamed Chahine Ghanem11 Sept 2026
PRIME-SVR: Physics-infoRmed Implicit Multi-Echo Slice-to-Volume Reconstruction for Fetal T2 mappingarXiv:2607.20136physics.med-phBusra Bulut, Maik Dannecker, Thomas Sanchez +211 Sept 2026
DexterSQL: Deep Schema Exploration and Rule-based Correction for Text-to-SQL GenerationarXiv:2608.11889cs.DBAnik Pramanik, Murat Kantarcioglu, Vincent Oria +111 Sept 2026
Left-Branching Transformers Excel at Right-Branching Languages: Data Shapes Word Order Preferences in Language ModelsarXiv:2608.15129cs.CLVarvara Arzt, Allan Hanbury, Terra Blevins11 Sept 2026
Chameleon: An Adaptive AI-Driven Honeypot Architecture Using Threat-Calibrated Particle Swarm Optimization and Semantic Deception Rapidly-Exploring Random TreesarXiv:2608.15407cs.CRRohit Swami, Tushar Singh, Akash Warde +111 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.