paper-with-me

홈 › Papers

Beyond Voxel 3D Editing: Learning from 3D Masks and Self-Constructed Data

2026-04-15 · Yizhao Xu, Hongyuan Zhu, Caiyun Liu, Tianfu Wang, Keyu Chen, Sicheng Xu, Jiaolong Yang, Nicholas Jing Yuan, Qi Zhang arxiv

3D editing refers to the ability to apply local or global modifications to 3D assets. Effective 3D editing requires maintaining semantic consistency by performing localized changes according to prompts, while also preserving local invariance so that unchanged regions remain consistent with the original. However, existing approaches have significant limitations: multi-view editing methods incur losses when projecting back to 3D, while voxel-based editing is constrained in both the regions that can be modified and the scale of modifications. Moreover, the lack of sufficiently large editing datasets for training and evaluation remains a challenge. To address these challenges, we propose a Beyond Voxel 3D Editing (BVE) framework with a self-constructed large-scale dataset specifically tailored for 3D editing. Building upon this dataset, our model enhances a foundational image-to-3D generative architecture with lightweight, trainable modules, enabling efficient injection of textual semantics without the need for expensive full-model retraining. Furthermore, we introduce an annotation-free 3D masking strategy to preserve local invariance, maintaining the integrity of unchanged regions during editing. Extensive experiments demonstrate that BVE achieves superior performance in generating high-quality, text-aligned 3D assets, while faithfully retaining the visual characteristics of the original input.

📄 PDF Abstract BibTeX arXiv:2604.13688

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

NANO3D: A Training-Free Approach for Efficient 3D Editing Without Masks

2025-10-16 · Junliang Ye, Shenghao Xie, Ruowen Zhao, Zhengyi Wang 외 arxiv

3D object editing is essential for interactive content creation in gaming, animation, and robotics, yet current approaches remain inefficient, inconsistent, and often fail to preserve unedited regions. Most methods rely …

3D Object Editing

DVD: Discrete Voxel Diffusion for 3D Generation and Editing

2026-05-08 · Zhengrui Xiang, Jiaqi Wu, Fupeng Sun, Heliang Zheng 외 arxiv

We introduce Discrete Voxel Diffusion (DVD), a discrete diffusion framework to generate, assess, and edit sparse voxels for SLat (Structured LATent) based 3D generative pipelines. Although discrete diffusion has not gene…

3D Generation

FreeMask: Rethinking the Importance of Attention Masks for Zero-Shot Video Editing

2024-09-30 · Lingling Cai, Kang Zhao, Hangjie Yuan, Yingya Zhang 외

Text-to-video diffusion models have made remarkable advancements. Driven by their ability to generate temporally coherent videos, research on zero-shot video editing using these fundamental models has expanded rapidly. T…

DenoisingVideo Editing

A Soft STAPLE Algorithm Combined with Anatomical Knowledge

2019-10-26 · Eytan Kats, Jacob Goldberger, Hayit Greenspan

Supervised machine learning algorithms, especially in the medical domain, are affected by considerable ambiguity in expert markings. In this study we address the case where the experts' opinion is obtained as a distribut…

Segmentation

Beyond Simple Edits: X-Planner for Complex Instruction-Based Image Editing

2025-07-07 · Chun-Hsiao Yeh, Yilin Wang, Nanxuan Zhao, Richard Zhang 외 arxiv

Recent diffusion-based image editing methods have significantly advanced text-guided tasks but often struggle to interpret complex, indirect instructions. Moreover, current models frequently suffer from poor identity pre…

Image Editing