| Reconstruction of a 3D wireframe from a single line drawing via generative depth estimationarXiv:2604.13549 | Elton Cao, Hod Lipson | 11 Sept 2026 | cs.CV | — | 89 |
| TextAlign: Preference Alignment for Text Rendering with Hierarchical RewardsarXiv:2605.19320 | Mingxuan Cui, Jingpu Yang, Fengxian Ji +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Artic-O: End-to-End Articulated Object Reconstruction via Latent Geometry LearningarXiv:2606.21938 | Xuyang Wang, Zhenyu Li, Jian Ding +2 | 11 Sept 2026 | cs.CV | — | 89 |
| ABACUS: Adapting Unified Foundation Model for Bridging Image Count Understanding and GenerationarXiv:2606.23835 | Anindya Mondal, Sauradip Nag, Anjan Dutta | 11 Sept 2026 | cs.CV | — | 89 |
| Does YOLO26 Truly Offer Advantages Over Its Predecessors for Edge Deployment? A Benchmark Study in AquaculturearXiv:2607.09835 | Rakesh Ranjan, Gajanan S. Kothawade, Kata Sharrer +2 | 11 Sept 2026 | cs.CV | — | 89 |
| HeteroPROMPT: A Real-time and Privacy-Preserving Heterogeneous Collaborative Perception FrameworkarXiv:2607.26283 | Armin Maleki, Hayder Radha | 11 Sept 2026 | cs.CV | — | 89 |
| SparSTAR: Sparse Attention for SpaceTime AutoRegressive Video SynthesisarXiv:2608.10519 | Jongbeom Lee, Hyunwoo Yu, Jincheol Yang +2 | 11 Sept 2026 | cs.CV | — | 89 |
| StreamTTT: Reconciling Real-Time Perception and Long-Term Memory in Streaming VLMsarXiv:2608.13416 | Joya Chen, Zeyun Zhong, Mike Zheng Shou | 11 Sept 2026 | cs.CV | — | 89 |
| Routing Before Looking: Query-Adaptive Evidence Acquisition for Long-form Video UnderstandingarXiv:2608.20805 | Tianyue Wang, Xuying Wu, Yuxiang Ma +2 | 11 Sept 2026 | cs.CV | — | 89 |
| MRI-based Deep Radiomic Phenotyping of Neuromuscular Disorders: A Topology-driven CharacterizationarXiv:2608.24415 | Martyna \.Zur, {\L}ukasz Pi\'orecki, Marek Socha +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Differentiable Jitter Correction using Deep Learning-based Image Quality Metric for Phase-Contrast Micro-CTarXiv:2608.27034 | Junan Chen, Yiting Jia, Joscha Maier +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Streaming4D: Accelerate 4D World Models via Block-wise Video Generation and Incremental ReconstructionarXiv:2609.00610 | Xiaoyan Liu, Jiaxin Liu, Kangrui Li +1 | 11 Sept 2026 | cs.CV | — | 89 |
| Design and Implementation of a Kalman Filter-Infused Algorithm for Tilt EstimationarXiv:2609.00730 | Yuehan Ma, Hongji Dai | 11 Sept 2026 | cs.CV | — | 89 |
| Persistent Identity Preservation in Generative Image Models: A Benchmark and Evaluation SystemarXiv:2609.04151 | Mengwei Ren, Xuaner Zhang, Zhihao Xia | 11 Sept 2026 | cs.CV | — | 89 |
| FreeTransformSR: Efficient Lightweight Image Super-Resolution via Free Low-Rank Learnable TransformarXiv:2609.05912 | Hongji Li, Yunhui Li | 11 Sept 2026 | cs.CV | — | 89 |
| FujinSplat: Seeing Through Smoke with RAW-Domain Gaussian SplattingarXiv:2609.06017 | Gengjia Chang, Ziteng Cui, Shuhong Liu | 11 Sept 2026 | cs.CV | — | 89 |
| RAIDAL: Redundancy-Aware Information Density Active Learning for CTC-Based Continuous Sign Language RecognitionarXiv:2609.06843 | Rafael A. Diniz Augusto, Gabriel L. Oliveira, Erickson R. Nascimento | 11 Sept 2026 | cs.CV | — | 89 |
| CGSM: Concept-Guided Segmentation Model for Precise Pulmonary Lesion DelineationarXiv:2609.07004 | Changheng Lin, Wenjie Zhang, Yushan Lu +2 | 11 Sept 2026 | cs.CV | — | 89 |
| Ambient @ EgoLongQA 2026: Distilling Long-Video perception into a Sub-2B ModelarXiv:2609.07154 | Logesh Kumar Umapathi | 11 Sept 2026 | cs.CV | — | 89 |
| From Few-Shot Segmentation to Clinician-in-the-Loop Medical Image AnalysisarXiv:2609.10001 | Yazhou Zhu | 11 Sept 2026 | cs.CV | — | 89 |
| 3rd Place Solution to Human Motion Challenges in Real-World and Clinical Settings (MoCha) @ECCV2026: Language-Aligned Motion Representations for Domain-Generalizable UPDRS-Gait Severity EstimationarXiv:2609.10187 | Soojie Kim, Muhammad Munsif, Minkyung Kim +1 | 11 Sept 2026 | cs.CV | — | 89 |
| SegKAN: High-Resolution Medical Image Segmentation with Long-Distance DependenciesarXiv:2412.19990 | Shengbo Tan, Rundong Xue, Shipeng Luo +2 | 11 Sept 2026 | eess.IV | — | 89 |
| PathoHR: Breast Cancer Survival Prediction on High-Resolution Pathological ImagesarXiv:2503.17970 | Yang Luo, Shiru Wang, Jun Liu +2 | 11 Sept 2026 | eess.IV | — | 89 |
| FastMap: Real-Time Semantic Map Completion via Bitwise Masked ModelingarXiv:2506.07350 | Yijie Deng, Shuaihang Yuan, Congcong Wen +2 | 11 Sept 2026 | cs.RO | — | 89 |
| Prompting with Sign Parameters for Low-resource Sign Language Instruction GenerationarXiv:2508.16076 | Md Tariquzzaman, Md Farhan Ishmam, Saiyma Sittul Muna +2 | 11 Sept 2026 | cs.HC | — | 89 |
| DCReg: Decoupled Characterization for Efficient Degenerate LiDAR RegistrationarXiv:2509.06285 | Xiangcheng Hu, Xieyuanli Chen, Mingkai Jia +2 | 11 Sept 2026 | cs.RO | — | 89 |
| DefVINS: Visual-Inertial Odometry for Deformable ScenesarXiv:2601.00702 | Samuel Cerezo, Javier Civera | 11 Sept 2026 | cs.RO | — | 89 |
| Measuring Browser Webcam Gaze Honestly: A Capture-Clock Methodology and Open Reference ImplementationarXiv:2608.11566 | Chi-Sheng Chen, Gabriel A. Brat | 11 Sept 2026 | cs.HC | — | 89 |
| Diagnosing and Dynamically Filtering Occupancy World Models for Active MappingarXiv:2609.06820 | Jiahui Zhang, Gongbo Liang, Yu Zhang | 11 Sept 2026 | cs.RO | — | 89 |
| TBR: Transport-Based Rendering with Deposition Strokes for Inverse GraphicsarXiv:2609.08722 | Tianqi Liu, Yushan Han, Hang Liu | 11 Sept 2026 | cs.GR | — | 89 |
| Data-Driven Risk Fields for Safer End-to-End Autonomous DrivingarXiv:2609.10377 | Yuanxin Tian, Zhiyuan Liu, Jinhao Li +2 | 11 Sept 2026 | cs.RO | — | 89 |