| No Free Checker: A Survey of Verifiers for Robot PoliciesarXiv:2609.09250 | Yang Wan, Xihang Yue, Zhirui Liu +2 | 11 Sept 2026 | cs.RO | — | 89 |
| VANTAGE-Bench: Evaluating the Infrastructure AI Gap in Vision-Language ModelsarXiv:2609.09396 | Zaid Pervaiz Bhat, Nimra Nayyar, Arihant Jain +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Myocardial Strain Drift Correction in Deep Learning Based Ultrasound TrackingarXiv:2609.09577 | Thierry Judge, Nicolas Duchateau, Andreas {\O}stvik +2 | 11 Sept 2026 | eess.IV | — | 89 |
| RouteBridge: Reliability-Routed Bidirectional Distillation Between Neural Radiance Fields and 3D Gaussian SplattingarXiv:2609.09606 | YuanHang Wang, Xin Cao | 11 Sept 2026 | cs.CV | — | 89 |
| Hyperbolic Geometry for Open-World Object Detection in Remote Sensing ImageryarXiv:2609.09626 | Wuzhou Li, Jiawei Zhou, Shenghang Wang +1 | 11 Sept 2026 | cs.CV | — | 89 |
| Distilling Image Prototypes for Guided Test-Time AdaptationarXiv:2609.09737 | Liwen Wang, Xingbo Dong, Iman Yi Liao +2 | 11 Sept 2026 | cs.CV | — | 89 |
| LogiScope-VQA: Benchmarking Vision-Language Models for Logistics Hazard Identification in Industrial ScenariosarXiv:2609.09790 | Hanjing Zhou, Mingze Yin, Ying Lian +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Albedo Estimation via Latent Bridge MatchingarXiv:2609.09884 | Carme Corbi, David Serrano-Lozano, Javier Vazquez-Corral +1 | 11 Sept 2026 | cs.CV | — | 89 |
| Strangers to Themselves: What Language Models Say About Themselves Is GenericarXiv:2609.09899 | Phil Blandfort, Urja Pawar | 11 Sept 2026 | cs.LG | — | 89 |
| FlowCPO: A Unified Divergence View of Preference Alignment for Flow ModelsarXiv:2609.09905 | Yansen Han, Shengyi Liao, Peng Sun +2 | 11 Sept 2026 | stat.ML | — | 89 |
| What Makes Adversarial Examples Transfer Across Deepfake Detectors?arXiv:2609.10002 | Rafael M. Mamede, Pedro C. Neto, Ana F. Sequeira | 11 Sept 2026 | cs.CV | — | 89 |
| Elastoformer: Enabling Dynamic Adaptivity via Elastic Model TransformationarXiv:2609.10018 | Sudaksh Kalra, Dolly Sapra | 11 Sept 2026 | cs.CV | — | 89 |
| A statistical approach to bias in zero-shot learning: the lens of handwriting recognitionarXiv:2609.10084 | Clarence Chew, Gim Siang Chia, Sukalpa Chanda +2 | 11 Sept 2026 | stat.ML | — | 89 |
| SA-Profile: Automated Sulcus Angle Profiling from Super-Resolution MRIarXiv:2609.10125 | Michael Wehrli, Leo Widmer, Edwin Li +2 | 11 Sept 2026 | cs.CV | — | 89 |
| One Loop, Two Gains: Can Active Learning win the Lottery for Free?arXiv:2609.10311 | Benedikt Tscheschner, Eduardo Veas, Marc Masana | 11 Sept 2026 | cs.LG | — | 89 |
| Beyond One-Size-Fits-All: Sample-Adaptive Strategy Routing for Vision Token Pruning in MLLMsarXiv:2609.10346 | Haiji Liang, Pengfei Zhou, Zhenglin Wan +2 | 11 Sept 2026 | cs.CV | — | 89 |
| PACE: Perceived-Latency-Aware Cascading Service Routing and Filler Control for QoE-Efficient Retrieval-Augmented Dialogue ServingarXiv:2609.10372 | Lin Huang, Yujuan Tan, Weisheng Li +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Semigroup-JEPA: Latent Dynamics Consistency for Zero-Shot Physics GeneralizationarXiv:2609.10464 | Andy Zeyi Liu, Haoran Sun, Lucas Baker +2 | 11 Sept 2026 | cs.LG | — | 89 |
| Show-Harness: Just a VLM Agent Can Play RobotsarXiv:2609.10522 | Yanzhe Chen, Zechen Bai, Zhijun Cao +2 | 11 Sept 2026 | cs.RO | — | 89 |
| Zero-shot World Models Are Developmentally Efficient LearnersarXiv:2604.10333 | Khai Loong Aw, Klemen Kotar, Wanhee Lee +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Learning to Predict Middle-Layer Attention in MLLMs for Visual Token PruningarXiv:2608.06411 | Yuyao Sun, Tao Deng, Shuang Li +2 | 11 Sept 2026 | cs.AI | — | 89 |
| RevalExo: A Functional Daily-Activity Benchmark for Inertial and Visual Locomotion Mode Recognition in Older Adults and Clinical CohortsarXiv:2609.08090 | Diwas Lamsal, Juha Carlon, Reinhard Claeys +2 | 11 Sept 2026 | cs.AI | — | 89 |
| Synergistic Vision-Language Reinforcement Enables Scalable On-Demand Analysis across Diverse Clinical TasksarXiv:2505.03380 | Haonan Wang, Jiaji Mao, Lehan Wang +2 | 11 Sept 2026 | cs.CV | — | 89 |
| SloMoDeblur: A Large-Scale Smartphone Image Deblurring DatasetarXiv:2506.19445 | Syed Mumtahin Mahmud, Mahdi Mohd Hossain Noki, Prothito Shovon Majumder +2 | 11 Sept 2026 | cs.CV | — | 89 |
| RAU: Reference-based Anatomical Understanding with Vision Language ModelsarXiv:2509.22404 | Yiwei Li, Yikang Liu, Jiaqi Guo +2 | 11 Sept 2026 | cs.CV | — | 89 |
| FiberTune: Preserving Action-Fiber Visual Residuals in Vision-Language-Action Fine-TuningarXiv:2606.08653 | Haihao Lin, Xiangsheng Huang, Xiao Yang +2 | 11 Sept 2026 | cs.CV | — | 89 |
| PSCT-Net: Geometry-Aware Pediatric Skull CT Reconstruction via Differentiable Back-Projection and Attention-Guided RefinementarXiv:2606.19867 | Dong Yeong Kim, Jaewon Choi, Youmin Shin +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Phase-Aware Spatial-Frequency Fusion for Few-Shot Fine-Grained Image ClassificationarXiv:2609.03829 | Ruiling Liu, Linyue Zhang, Wenyi Zeng +2 | 11 Sept 2026 | cs.CV | — | 89 |
| When Does a Laugh Begin? Structured Annotator Disagreement in Temporal Laughter LocalizationarXiv:2609.06646 | Eyal Hanania, Daniel Arkushin, Naveh Ayal +2 | 11 Sept 2026 | cs.CV | — | 89 |
| SAFER-Activities: A Dataset for Smart Assessment of Fall Events and Routine ActivitiesarXiv:2609.08038 | Diwas Lamsal, Pramod Wickramatilake, Jednipat Moonrinta +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Hi-FLoop: Hierarchical State-Feedback Loops for Multi-Timescale World ModelingarXiv:2609.08796 | Rx Fan, Z Han | 11 Sept 2026 | cs.CV | — | 89 |
| Rethinking Handwritten Character RecognitionarXiv:2609.10572 | Ranjit Raut, Aarav Subedi, Ashim Shrestha | 11 Sept 2026 | cs.CV | — | 89 |
| AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation FlowarXiv:2609.10723 | Junran Wang, Zehao Jin, Tianyu Luan +1 | 11 Sept 2026 | cs.CV | — | 89 |
| MHE-Former: Multi-Hypothesis Transformers via Entropy Maximization for 3D Mesh RecoveryarXiv:2609.10743 | Boshu Jia, Rongyu Chen, Linlin Yang +2 | 11 Sept 2026 | cs.CV | — | 89 |
| GRADE: Single-Frame Generative Radar Depth Estimation Under Visual DegradationarXiv:2609.10756 | Bin Zhao, Patrick Chiou, Nakul Garg | 11 Sept 2026 | cs.CV | — | 89 |
| Shedding Light: A Benchmark for Evaluating Lighting Understanding in Generative Image ModelsarXiv:2609.10787 | Justine Giroux, Jack Oliver Hilliard, Yannick Hold-Geoffroy +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Two-Parameter Flow Map Learning for Continuous-Time Diffeomorphic Image RegistrationarXiv:2609.10789 | Mohammadjavad Matinkia, Nilanjan Ray | 11 Sept 2026 | cs.CV | — | 89 |
| TrajFusionNet+: Transformer-Based Prediction of Pedestrian Crossing Intention via Fusion of Trajectory Representations and Scene GraphsarXiv:2609.10806 | Fran\c{c}ois G. Landry, Moulay A. Akhloufi | 11 Sept 2026 | cs.CV | — | 89 |
| Overpainting: Localized Context-aware Diffusion Image EditingarXiv:2609.10811 | Sam Sartor, Iliyan Georgiev, Michael Fischer +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Evaluation of Vision-Language Models Across Diverse Coastal EnvironmentsarXiv:2609.10855 | Seth Knoop, Chad R. Samuelson, Gabriel R. Slade +2 | 11 Sept 2026 | cs.CV | — | 89 |