Skip to content
AI Atlas
Research

Papers

Publications linked to models, labs and benchmarks. Authors, venues and abstracts come from arXiv and publisher pages; model links come from model cards citing the paper.

1,245 papers

Papers
TitleAuthorsOrganizationPublishedIntroduces Models (and artifacts) whose model card or documentation cites this paper — inbound described_by relations.DatasetsBenchmarksCode
FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow EstimationarXiv:2609.11486cs.CVVladislav Bargatin, Alexander Yakovenko, Khaled Abud +111 Sept 2026
Recursive Code World Models: Building Complex Worlds through Recursive Scene ProgramsarXiv:2609.11499cs.CVZhiqi Li, Yuxuan Liao, Bo Zhu11 Sept 2026
UBone3D: Physics-Rectified Conditional Flow Matching for Anatomical 3D Shape Completion from UltrasoundarXiv:2609.11506cs.CVWeiying Chen, Yuchong Gao, Siyuan Li +211 Sept 2026
Harnessing Intrinsic Subject-Aware Attention for Controllable Multi-Subject Video GenerationarXiv:2609.11507cs.CVNiange Yu, Ye Tian, Biaolong Chen +211 Sept 2026
Prototype Matters: Modality-unified Prototype Self-distillation for Unsupervised Visible-infrared Person Re-identificationarXiv:2609.11514cs.CVMenglin Wang, Xiaojin Gong11 Sept 2026
LoopVAE: Recurrent Depth Across Scales for Visual TokenizationarXiv:2609.11516cs.CVZhiying Lu11 Sept 2026
World in World: Explore the World with World ModelsarXiv:2609.11548cs.CVChenxi Song, Yanming Yang, Chi Zhang11 Sept 2026
OmniKVQuant: KV Cache Quantization for Omni-LLMsarXiv:2609.11582cs.CVSuho Yoo, Hyunjong Ok, Jongmin Choi +211 Sept 2026
MMGait: Benchmarking and Unifying Gait Recognition across Heterogeneous ModalitiesarXiv:2609.11601cs.CVSaihui Hou, Chenye Wang, Qingyuan Cai +211 Sept 2026
LangStreet: Persistent Language Fields for Anchor-Decoded Street GaussiansarXiv:2609.11616cs.CVRunyi Yang, Deheng Zhang, Xiaoye Wang +211 Sept 2026
Self-Supervised Cardiac Phase Detection via Single-Parameter Latent OrbitsarXiv:2609.11650cs.CVJohn Bonnici, Matthew Baugh, Aleksandra Kulbaka +211 Sept 2026
Single-Stream Multi-Feature Fusion with Temporal Robustness for Gait Emotion RecognitionarXiv:2609.11680cs.CVShirong Lyu, Silu Quan, Yixuan Ding +111 Sept 2026
Spectral Adapters for Segment Anything Model-based Segmentation of Colorectal Liver Metastases in Computed TomographyarXiv:2609.11703cs.CVRamtin Mojtahedi, Mohammad Hamghalam, Jacob J. Peoples +211 Sept 2026
MC-DeTra: Motion-Consistent Joint Object Detection and Socially-Aware Trajectory Forecasting in Bird's-Eye-View ImagesarXiv:2609.11717cs.CVVladislav Diuzhev, Dmitry Yudin11 Sept 2026
Revisiting Avatar-As-Image: High-Fidelity Registration is All You NeedarXiv:2609.11722cs.CVMargaret Kostyrko, Yuxuan Xue, Garvita Tiwari +111 Sept 2026
Guided Super-Resolution of Digital Elevation Models with Diffusion-Based Image GeneratorsarXiv:2609.11886cs.CVArmand Mihai Nicolicioiu, Dominik Narnhofer, Nando Metzger +211 Sept 2026
Caption-once, Frames-on-Demand: Visual-Need Routing for Budget-Aware Agentic Long Video UnderstandingarXiv:2609.11899cs.CVWeitong Cai, Hang Zhang, Yukai Huang +211 Sept 2026
SenseNova-U1.5: Towards Native Unified Visual IntelligencearXiv:2609.11929cs.CVHaiwen Diao, Jiahao Wang, Chenjing Ding +211 Sept 2026
Seamless Whole Slide Label-Free Virtual StainingarXiv:2609.10914eess.IVDou Hoon Kwark, Kianoush Falahkheirkhah, Ji-hun Oh +211 Sept 2026
IMLE-VLA: Fast Single-Step Action Generation for Vision-Language-Action PoliciesarXiv:2609.10915cs.ROKian Hosseinkhani (Simon Fraser University), Qinhe Peng (University of Pennsylvania), George Shramko (Simon Fraser University) +211 Sept 2026
Exponential Pixelating Integral transform with dual fractal features for enhanced chest X-ray abnormality detectionarXiv:2609.10988eess.IVNaveenraj Kamalakannan, Sri Ram Macharla, M Kanimozhi +111 Sept 2026
SegCol Challenge: Semantic Segmentation for Tools and Fold Edges in Colonoscopy dataarXiv:2412.16078cs.CVXinwei Ju, Rema Daher, Razvan Caramalau +211 Sept 2026
SSS: Semi-Supervised SAM-2 with Efficient Prompting for Medical Imaging SegmentationarXiv:2506.08949cs.CVHongjie Zhu, Xiwei Liu, Rundong Xue +211 Sept 2026
Dream4D: Lifting Camera-Controlled I2V towards Spatiotemporally Consistent 4D GenerationarXiv:2508.07769cs.CVXiaoyan Liu, Kangrui Li, Jiaxin Liu +211 Sept 2026
Adaptive Dual-Constrained Line Aggregation for Cross-Paradigm Line Segment DetectionarXiv:2508.19742cs.CVChenguang Liu, Chisheng Wang, Huilin Chen +211 Sept 2026

Author lists and categories are copied from the paper's own metadata (arXiv, publisher). Linked models, datasets, benchmarks and code come from stated relations only; a dash means no source stated one.