Papers
arxiv:2608.01492

GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization

Published on Aug 2
ยท Submitted by
YuhengLiu
on Aug 7
ยท lllab LL Lab
Authors:
,
,
,
,
,

Abstract

Selecting a complete 3D object from a reconstructed scene with minimal user effort is essential for practical scene editing and embodied interaction. Existing 3DGS-based methods either retrain the Gaussian representation to embed per-object labels, or build dense multi-view SAM observations, both requiring heavy computation and dense viewpoint coverage that is rarely available in practice. We present GaussianSelector, a training-free framework for interactive 3D object selection from sparse views and sparse scribble guidance. Operating directly on native Gaussian primitives, we coarsen dense Gaussians into geometrically coherent superpoints and construct a continuity-weighted graph using appearance and spatial cues. Sparse user scribbles are lifted into 3D via visibility-aware transmittance coverage, and selection is solved as a global graph-cut energy minimization that propagates sparse evidence to a complete 3D object. This design naturally supports multi-round refinement, where users iteratively correct the selection from additional viewpoints to progressively improve the result. Experiments demonstrate that GaussianSelector achieves competitive selection quality against state-of-the-art multi-view SAM-based methods, while requiring significantly fewer interaction views and substantially lower computational overhead. These properties make it well suited for human-in-the-loop 3D scene editing and 3D asset extraction in real-world deployment scenarios.

Community

Paper author Paper submitter

We built Photoshop's magic wand for 3D Gaussian Splatting. โœจ

Meet ๐†๐š๐ฎ๐ฌ๐ฌ๐ข๐š๐ง๐’๐ž๐ฅ๐ž๐œ๐ญ๐จ๐ซ ๐Ÿ–Œ๏ธ โ€” scribble on one view, and the whole 3D object pops out of the scene. Same gesture you'd use to select a flower in a photo, except the canvas is in 3D space and what you get back is a real 3D asset.

Forget about renting GPU servers to deploy heavy 3D segmentation models. That's a phone doing it, live ๐Ÿ“ฑ๐Ÿ‘‡

This is an automated message from the Librarian Bot. I found the following papers similar to this paper.

The following papers were recommended by the Semantic Scholar API

Please give a thumbs up to this comment if you found it helpful!

If you want recommendations for any Paper on Hugging Face checkout this Space

You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend

Sign up or log in to comment

Get this paper in your agent:

hf papers read 2608.01492
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 0

No model linking this paper

Cite arxiv.org/abs/2608.01492 in a model README.md to link it from this page.

Datasets citing this paper 0

No dataset linking this paper

Cite arxiv.org/abs/2608.01492 in a dataset README.md to link it from this page.

Spaces citing this paper 0

No Space linking this paper

Cite arxiv.org/abs/2608.01492 in a Space README.md to link it from this page.

Collections including this paper 1