paper-with-me

Papers

Point'n Move: Interactive Scene Object Manipulation on Gaussian Splatting Radiance Fields

2023-11-28 · Jiajun Huang, Hongchuan Yu

We propose Point'n Move, a method that achieves interactive scene object manipulation with exposed region inpainting. Interactivity here further comes from intuitive object selection and real-time editing. To achieve this, we adopt Gaussian Splatting Radiance Field as the scene representation and fully leverage its explicit nature and speed advantage. Its explicit representation formulation allows us to devise a 2D prompt points to 3D mask dual-stage self-prompting segmentation algorithm, perform mask refinement and merging, minimize change as well as provide good initialization for scene inpainting and perform editing in real-time without per-editing training, all leads to superior quality and performance. We test our method by performing editing on both forward-facing and 360 scenes. We also compare our method against existing scene object removal methods, showing superior quality despite being more capable and having a speed advantage.

📄 PDF Abstract BibTeX arXiv:2311.16737

Code (0)

등록된 구현이 없습니다.

Tasks

Object

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
Inpainting Train a convolutional neural network to generate the contents of an arbitrary image region conditioned on its surroundings.

Similar Papers 제목 키워드 기반

Articulated Object Manipulation using Online Axis Estimation with SAM2-Based Tracking

2024-09-24 · Xi Wang, Tianxing Chen, Qiaojun Yu, Tianling Xu 외

Articulated object manipulation requires precise object interaction, where the object's axis must be carefully considered. Previous research employed interactive perception for manipulating articulated objects, but typic…

Object

WorldCraft: From Camera Navigation to Object Manipulation in Interactive Video World Models

2026-05-24 · Bohai Gu, Taiyi Wu, Yueyang Yuan, Jian Liu 외 arxiv

Recent video-based world models have made pixel-space environments interactive at the camera level: users can navigate viewpoints while the model generates coherent visual continuations. Yet their action spaces remain in…

Neural Point Catacaustics for Novel-View Synthesis of Reflections

2023-01-03 · Georgios Kopanas, Thomas Leimkühler, Gilles Rainer, Clément Jambon 외

View-dependent effects such as reflections pose a substantial challenge for image-based and neural rendering algorithms. Above all, curved reflectors are particularly hard, as they lead to highly non-linear reflection fl…

Neural RenderingNovel View Synthesis

CoReLIN: Constraint-based Reasoning for Zero-shot Lifelong Interactive Navigation

2026-02-23 · Apoorva Vashisth, Manav Kulshrestha, Pranav Bakshi, Damon Conover 외 arxiv

Robot navigation typically assumes an obstacle-free path exists between start and goal. In real environments, however, clutter may block all routes. We introduce Lifelong Interactive Navigation, where a mobile robot with…

Robot Navigation

Bootstrapping Robotic Ecological Perception from a Limited Set of Hypotheses Through Interactive Perception

2019-01-30 · Léni K. Le Goff, Ghanim Mukhtar, Alexandre Coninx, Stéphane Doncieux

To solve its task, a robot needs to have the ability to interpret its perceptions. In vision, this interpretation is particularly difficult and relies on the understanding of the structure of the scene, at least to the e…