Skip to content
AI Atlas
PaperActive

SAMV-DUSt3R: Instance-Centric 3D Scene Decoupling from Sparse Multi-Views

arxiv.org/abs/2609.11279

Updated 50 min ago · first seen 11 Sept 2026

paper_01M294H1ZWV3460K1D00QYC5N1

Published
11 Sept 2026
T1 · 50 min ago
arXiv
2609.11279
T1 · 50 min ago
Category
cs.CV
T1 · 50 min ago

Abstract

With the rising demand to decouple objects from 3D scenes, we propose SAMV-DUSt3R, an end-to-end model that injects SAM2 2D masks into MV-DUSt3R reconstruction. A Cross Flow Mask Block uses these masks to steer the network toward the target instance, jointly improving shape accuracy and achieving object-level disentanglement without multi-stage pipelines. To ensure reconstruction stability, a lightweight Spatial RankGNN selects the optimal reference view with a selection accuracy of 73.5\%. Extensive experiments demonstrate that our method boosts average reconstruction precision by 11\% across various metrics compared to state-of-the-art baselines. These results reveal a strong instance-disentanglement capability and clear benefits for driving, robotics, AR/VR, and heritage digitisation.

Authors 5

Langxu Zhao, Zuan Gu, Yingdan Zhang, Pengfei Zhao, Tianhan Gao

Specification

Official page

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Arxiv announce type
new

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

arXiv id
2609.11279

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Categories
cs.CV

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

PDF

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Primary category
cs.CV

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Published
11 Sept 2026

Source:arXiv (Atom API + RSS)T1observed 50 min agohigh

Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →

Provenance

Attributed facts

9

Source tiers

T19

Freshest observation

50 min ago

Conflicts

None