| FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow EstimationarXiv:2609.11486cs.CV | Vladislav Bargatin, Alexander Yakovenko, Khaled Abud +1 | — | 11 Sept 2026 | — | — | — | — |
| Recursive Code World Models: Building Complex Worlds through Recursive Scene ProgramsarXiv:2609.11499cs.CV | Zhiqi Li, Yuxuan Liao, Bo Zhu | — | 11 Sept 2026 | — | — | — | — |
| UBone3D: Physics-Rectified Conditional Flow Matching for Anatomical 3D Shape Completion from UltrasoundarXiv:2609.11506cs.CV | Weiying Chen, Yuchong Gao, Siyuan Li +2 | — | 11 Sept 2026 | — | — | — | — |
| Harnessing Intrinsic Subject-Aware Attention for Controllable Multi-Subject Video GenerationarXiv:2609.11507cs.CV | Niange Yu, Ye Tian, Biaolong Chen +2 | — | 11 Sept 2026 | — | — | — | — |
| Prototype Matters: Modality-unified Prototype Self-distillation for Unsupervised Visible-infrared Person Re-identificationarXiv:2609.11514cs.CV | Menglin Wang, Xiaojin Gong | — | 11 Sept 2026 | — | — | — | — |
| LoopVAE: Recurrent Depth Across Scales for Visual TokenizationarXiv:2609.11516cs.CV | Zhiying Lu | — | 11 Sept 2026 | — | — | — | — |
| World in World: Explore the World with World ModelsarXiv:2609.11548cs.CV | Chenxi Song, Yanming Yang, Chi Zhang | — | 11 Sept 2026 | — | — | — | — |
| OmniKVQuant: KV Cache Quantization for Omni-LLMsarXiv:2609.11582cs.CV | Suho Yoo, Hyunjong Ok, Jongmin Choi +2 | — | 11 Sept 2026 | — | — | — | — |
| MMGait: Benchmarking and Unifying Gait Recognition across Heterogeneous ModalitiesarXiv:2609.11601cs.CV | Saihui Hou, Chenye Wang, Qingyuan Cai +2 | — | 11 Sept 2026 | — | — | — | — |
| LangStreet: Persistent Language Fields for Anchor-Decoded Street GaussiansarXiv:2609.11616cs.CV | Runyi Yang, Deheng Zhang, Xiaoye Wang +2 | — | 11 Sept 2026 | — | — | — | — |
| Self-Supervised Cardiac Phase Detection via Single-Parameter Latent OrbitsarXiv:2609.11650cs.CV | John Bonnici, Matthew Baugh, Aleksandra Kulbaka +2 | — | 11 Sept 2026 | — | — | — | — |
| Single-Stream Multi-Feature Fusion with Temporal Robustness for Gait Emotion RecognitionarXiv:2609.11680cs.CV | Shirong Lyu, Silu Quan, Yixuan Ding +1 | — | 11 Sept 2026 | — | — | — | — |
| Spectral Adapters for Segment Anything Model-based Segmentation of Colorectal Liver Metastases in Computed TomographyarXiv:2609.11703cs.CV | Ramtin Mojtahedi, Mohammad Hamghalam, Jacob J. Peoples +2 | — | 11 Sept 2026 | — | — | — | — |
| MC-DeTra: Motion-Consistent Joint Object Detection and Socially-Aware Trajectory Forecasting in Bird's-Eye-View ImagesarXiv:2609.11717cs.CV | Vladislav Diuzhev, Dmitry Yudin | — | 11 Sept 2026 | — | — | — | — |
| Revisiting Avatar-As-Image: High-Fidelity Registration is All You NeedarXiv:2609.11722cs.CV | Margaret Kostyrko, Yuxuan Xue, Garvita Tiwari +1 | — | 11 Sept 2026 | — | — | — | — |
| Guided Super-Resolution of Digital Elevation Models with Diffusion-Based Image GeneratorsarXiv:2609.11886cs.CV | Armand Mihai Nicolicioiu, Dominik Narnhofer, Nando Metzger +2 | — | 11 Sept 2026 | — | — | — | — |
| Caption-once, Frames-on-Demand: Visual-Need Routing for Budget-Aware Agentic Long Video UnderstandingarXiv:2609.11899cs.CV | Weitong Cai, Hang Zhang, Yukai Huang +2 | — | 11 Sept 2026 | — | — | — | — |
| SenseNova-U1.5: Towards Native Unified Visual IntelligencearXiv:2609.11929cs.CV | Haiwen Diao, Jiahao Wang, Chenjing Ding +2 | — | 11 Sept 2026 | — | — | — | — |
| Seamless Whole Slide Label-Free Virtual StainingarXiv:2609.10914eess.IV | Dou Hoon Kwark, Kianoush Falahkheirkhah, Ji-hun Oh +2 | — | 11 Sept 2026 | — | — | — | — |
| IMLE-VLA: Fast Single-Step Action Generation for Vision-Language-Action PoliciesarXiv:2609.10915cs.RO | Kian Hosseinkhani (Simon Fraser University), Qinhe Peng (University of Pennsylvania), George Shramko (Simon Fraser University) +2 | — | 11 Sept 2026 | — | — | — | — |
| Exponential Pixelating Integral transform with dual fractal features for enhanced chest X-ray abnormality detectionarXiv:2609.10988eess.IV | Naveenraj Kamalakannan, Sri Ram Macharla, M Kanimozhi +1 | — | 11 Sept 2026 | — | — | — | — |
| SegCol Challenge: Semantic Segmentation for Tools and Fold Edges in Colonoscopy dataarXiv:2412.16078cs.CV | Xinwei Ju, Rema Daher, Razvan Caramalau +2 | — | 11 Sept 2026 | — | — | — | — |
| SSS: Semi-Supervised SAM-2 with Efficient Prompting for Medical Imaging SegmentationarXiv:2506.08949cs.CV | Hongjie Zhu, Xiwei Liu, Rundong Xue +2 | — | 11 Sept 2026 | — | — | — | — |
| Dream4D: Lifting Camera-Controlled I2V towards Spatiotemporally Consistent 4D GenerationarXiv:2508.07769cs.CV | Xiaoyan Liu, Kangrui Li, Jiaxin Liu +2 | — | 11 Sept 2026 | — | — | — | — |
| Adaptive Dual-Constrained Line Aggregation for Cross-Paradigm Line Segment DetectionarXiv:2508.19742cs.CV | Chenguang Liu, Chisheng Wang, Huilin Chen +2 | — | 11 Sept 2026 | — | — | — | — |