Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages.

191 total

Reset
Papers
TitleAuthorsPublishedCategoriesOrganization / venueQuality
No Free Checker: A Survey of Verifiers for Robot PoliciesarXiv:2609.09250Yang Wan, Xihang Yue, Zhirui Liu +211 Sept 2026cs.RO89
VANTAGE-Bench: Evaluating the Infrastructure AI Gap in Vision-Language ModelsarXiv:2609.09396Zaid Pervaiz Bhat, Nimra Nayyar, Arihant Jain +211 Sept 2026cs.CV89
Myocardial Strain Drift Correction in Deep Learning Based Ultrasound TrackingarXiv:2609.09577Thierry Judge, Nicolas Duchateau, Andreas {\O}stvik +211 Sept 2026eess.IV89
RouteBridge: Reliability-Routed Bidirectional Distillation Between Neural Radiance Fields and 3D Gaussian SplattingarXiv:2609.09606YuanHang Wang, Xin Cao11 Sept 2026cs.CV89
Hyperbolic Geometry for Open-World Object Detection in Remote Sensing ImageryarXiv:2609.09626Wuzhou Li, Jiawei Zhou, Shenghang Wang +111 Sept 2026cs.CV89
Distilling Image Prototypes for Guided Test-Time AdaptationarXiv:2609.09737Liwen Wang, Xingbo Dong, Iman Yi Liao +211 Sept 2026cs.CV89
LogiScope-VQA: Benchmarking Vision-Language Models for Logistics Hazard Identification in Industrial ScenariosarXiv:2609.09790Hanjing Zhou, Mingze Yin, Ying Lian +211 Sept 2026cs.CV89
Albedo Estimation via Latent Bridge MatchingarXiv:2609.09884Carme Corbi, David Serrano-Lozano, Javier Vazquez-Corral +111 Sept 2026cs.CV89
Strangers to Themselves: What Language Models Say About Themselves Is GenericarXiv:2609.09899Phil Blandfort, Urja Pawar11 Sept 2026cs.LG89
FlowCPO: A Unified Divergence View of Preference Alignment for Flow ModelsarXiv:2609.09905Yansen Han, Shengyi Liao, Peng Sun +211 Sept 2026stat.ML89
What Makes Adversarial Examples Transfer Across Deepfake Detectors?arXiv:2609.10002Rafael M. Mamede, Pedro C. Neto, Ana F. Sequeira11 Sept 2026cs.CV89
Elastoformer: Enabling Dynamic Adaptivity via Elastic Model TransformationarXiv:2609.10018Sudaksh Kalra, Dolly Sapra11 Sept 2026cs.CV89
A statistical approach to bias in zero-shot learning: the lens of handwriting recognitionarXiv:2609.10084Clarence Chew, Gim Siang Chia, Sukalpa Chanda +211 Sept 2026stat.ML89
SA-Profile: Automated Sulcus Angle Profiling from Super-Resolution MRIarXiv:2609.10125Michael Wehrli, Leo Widmer, Edwin Li +211 Sept 2026cs.CV89
One Loop, Two Gains: Can Active Learning win the Lottery for Free?arXiv:2609.10311Benedikt Tscheschner, Eduardo Veas, Marc Masana11 Sept 2026cs.LG89
Beyond One-Size-Fits-All: Sample-Adaptive Strategy Routing for Vision Token Pruning in MLLMsarXiv:2609.10346Haiji Liang, Pengfei Zhou, Zhenglin Wan +211 Sept 2026cs.CV89
PACE: Perceived-Latency-Aware Cascading Service Routing and Filler Control for QoE-Efficient Retrieval-Augmented Dialogue ServingarXiv:2609.10372Lin Huang, Yujuan Tan, Weisheng Li +211 Sept 2026cs.CV89
Semigroup-JEPA: Latent Dynamics Consistency for Zero-Shot Physics GeneralizationarXiv:2609.10464Andy Zeyi Liu, Haoran Sun, Lucas Baker +211 Sept 2026cs.LG89
Show-Harness: Just a VLM Agent Can Play RobotsarXiv:2609.10522Yanzhe Chen, Zechen Bai, Zhijun Cao +211 Sept 2026cs.RO89
Zero-shot World Models Are Developmentally Efficient LearnersarXiv:2604.10333Khai Loong Aw, Klemen Kotar, Wanhee Lee +211 Sept 2026cs.AI89
Learning to Predict Middle-Layer Attention in MLLMs for Visual Token PruningarXiv:2608.06411Yuyao Sun, Tao Deng, Shuang Li +211 Sept 2026cs.AI89
RevalExo: A Functional Daily-Activity Benchmark for Inertial and Visual Locomotion Mode Recognition in Older Adults and Clinical CohortsarXiv:2609.08090Diwas Lamsal, Juha Carlon, Reinhard Claeys +211 Sept 2026cs.AI89
Synergistic Vision-Language Reinforcement Enables Scalable On-Demand Analysis across Diverse Clinical TasksarXiv:2505.03380Haonan Wang, Jiaji Mao, Lehan Wang +211 Sept 2026cs.CV89
SloMoDeblur: A Large-Scale Smartphone Image Deblurring DatasetarXiv:2506.19445Syed Mumtahin Mahmud, Mahdi Mohd Hossain Noki, Prothito Shovon Majumder +211 Sept 2026cs.CV89
RAU: Reference-based Anatomical Understanding with Vision Language ModelsarXiv:2509.22404Yiwei Li, Yikang Liu, Jiaqi Guo +211 Sept 2026cs.CV89
FiberTune: Preserving Action-Fiber Visual Residuals in Vision-Language-Action Fine-TuningarXiv:2606.08653Haihao Lin, Xiangsheng Huang, Xiao Yang +211 Sept 2026cs.CV89
PSCT-Net: Geometry-Aware Pediatric Skull CT Reconstruction via Differentiable Back-Projection and Attention-Guided RefinementarXiv:2606.19867Dong Yeong Kim, Jaewon Choi, Youmin Shin +211 Sept 2026cs.CV89
Phase-Aware Spatial-Frequency Fusion for Few-Shot Fine-Grained Image ClassificationarXiv:2609.03829Ruiling Liu, Linyue Zhang, Wenyi Zeng +211 Sept 2026cs.CV89
When Does a Laugh Begin? Structured Annotator Disagreement in Temporal Laughter LocalizationarXiv:2609.06646Eyal Hanania, Daniel Arkushin, Naveh Ayal +211 Sept 2026cs.CV89
SAFER-Activities: A Dataset for Smart Assessment of Fall Events and Routine ActivitiesarXiv:2609.08038Diwas Lamsal, Pramod Wickramatilake, Jednipat Moonrinta +211 Sept 2026cs.CV89
Hi-FLoop: Hierarchical State-Feedback Loops for Multi-Timescale World ModelingarXiv:2609.08796Rx Fan, Z Han11 Sept 2026cs.CV89
Rethinking Handwritten Character RecognitionarXiv:2609.10572Ranjit Raut, Aarav Subedi, Ashim Shrestha11 Sept 2026cs.CV89
AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation FlowarXiv:2609.10723Junran Wang, Zehao Jin, Tianyu Luan +111 Sept 2026cs.CV89
MHE-Former: Multi-Hypothesis Transformers via Entropy Maximization for 3D Mesh RecoveryarXiv:2609.10743Boshu Jia, Rongyu Chen, Linlin Yang +211 Sept 2026cs.CV89
GRADE: Single-Frame Generative Radar Depth Estimation Under Visual DegradationarXiv:2609.10756Bin Zhao, Patrick Chiou, Nakul Garg11 Sept 2026cs.CV89
Shedding Light: A Benchmark for Evaluating Lighting Understanding in Generative Image ModelsarXiv:2609.10787Justine Giroux, Jack Oliver Hilliard, Yannick Hold-Geoffroy +211 Sept 2026cs.CV89
Two-Parameter Flow Map Learning for Continuous-Time Diffeomorphic Image RegistrationarXiv:2609.10789Mohammadjavad Matinkia, Nilanjan Ray11 Sept 2026cs.CV89
TrajFusionNet+: Transformer-Based Prediction of Pedestrian Crossing Intention via Fusion of Trajectory Representations and Scene GraphsarXiv:2609.10806Fran\c{c}ois G. Landry, Moulay A. Akhloufi11 Sept 2026cs.CV89
Overpainting: Localized Context-aware Diffusion Image EditingarXiv:2609.10811Sam Sartor, Iliyan Georgiev, Michael Fischer +211 Sept 2026cs.CV89
Evaluation of Vision-Language Models Across Diverse Coastal EnvironmentsarXiv:2609.10855Seth Knoop, Chad R. Samuelson, Gabriel R. Slade +211 Sept 2026cs.CV89

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Each paper page shows the abstract, related models and every source snapshot.