paper-with-me

Papers

Structurally Disentangled Feature Fields Distillation for 3D Understanding and Editing

2025-02-20 · Yoel Levy, David Shavin, Itai Lang, Sagie Benaim

Recent work has demonstrated the ability to leverage or distill pre-trained 2D features obtained using large pre-trained 2D models into 3D features, enabling impressive 3D editing and understanding capabilities using only 2D supervision. Although impressive, models assume that 3D features are captured using a single feature field and often make a simplifying assumption that features are view-independent. In this work, we propose instead to capture 3D features using multiple disentangled feature fields that capture different structural components of 3D features involving view-dependent and view-independent components, which can be learned from 2D feature supervision only. Subsequently, each element can be controlled in isolation, enabling semantic and structural understanding and editing capabilities. For instance, using a user click, one can segment 3D features corresponding to a given object and then segment, edit, or remove their view-dependent (reflective) properties. We evaluate our approach on the task of 3D segmentation and demonstrate a set of novel understanding and editing tasks.

📄 PDF Abstract BibTeX arXiv:2502.14789

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding

2026-02-17 · Guile Wu, David Huang, Bingbing Liu, Dongfeng Bai arxiv

Existing 3D open-vocabulary scene understanding methods mostly emphasize distilling language features from 2D foundation models into 3D feature fields, but largely overlook the synergy among scene appearance, semantics, …

Scene Understanding

HandNeRF: Neural Radiance Fields for Animatable Interacting Hands

2023-03-24 · CVPR 2023 1 · Zhiyang Guo, Wengang Zhou, Min Wang, Li Li 외

We propose a novel framework to reconstruct accurate appearance and geometry with neural radiance fields (NeRF) for interacting hands, enabling the rendering of photo-realistic images and videos for gesture animation fro…

NeRF

Single-Domain Generalized Object Detection in Urban Scene via Cyclic-Disentangled Self-Distillation

2022-01-01 · CVPR 2022 1 · Aming Wu, Cheng Deng

In this paper, we are concerned with enhancing the generalization capability of object detectors. And we consider a realistic yet challenging scenario, namely Single-Domain Generalized Object Detection (Single-DGOD),…

Objectobject-detectionObject DetectionRobust Object Detection

Fast and Efficient: Mask Neural Fields for 3D Scene Segmentation

2024-07-01 · Zihan Gao, Lingling Li, Licheng Jiao, Fang Liu 외

Understanding 3D scenes is a crucial challenge in computer vision research with applications spanning multiple domains. Recent advancements in distilling 2D vision-language foundation models into neural fields, like NeRF…

3DGSNeRFScene Segmentation

Disentangled Generation and Aggregation for Robust Radiance Fields

2024-09-24 · Shihe Shen, Huachen Gao, Wangze Xu, Rui Peng 외

The utilization of the triplane-based radiance fields has gained attention in recent years due to its ability to effectively disentangle 3D scenes with a high-quality representation and low computation cost. A key requir…

NeRFNovel View Synthesis