paper-with-me

Papers

3DEnhancer: Consistent Multi-View Diffusion for 3D Enhancement

2024-12-24 · CVPR 2025 1 · Yihang Luo, Shangchen Zhou, Yushi Lan, Xingang Pan, Chen Change Loy

Despite advances in neural rendering, due to the scarcity of high-quality 3D datasets and the inherent limitations of multi-view diffusion models, view synthesis and 3D model generation are restricted to low resolutions with suboptimal multi-view consistency. In this study, we present a novel 3D enhancement pipeline, dubbed 3DEnhancer, which employs a multi-view latent diffusion model to enhance coarse 3D inputs while preserving multi-view consistency. Our method includes a pose-aware encoder and a diffusion-based denoiser to refine low-quality multi-view images, along with data augmentation and a multi-view attention module with epipolar aggregation to maintain consistent, high-quality 3D outputs across views. Unlike existing video-based approaches, our model supports seamless multi-view enhancement with improved coherence across diverse viewing angles. Extensive evaluations show that 3DEnhancer significantly outperforms existing methods, boosting both multi-view enhancement and per-instance 3D optimization tasks.

📄 PDF Abstract BibTeX arXiv:2412.18565

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationNeural Rendering

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Enhancing Nighttime UAV Tracking with Light Distribution Suppression

2024-09-25 · Liangliang Yao, Changhong Fu, Yiheng Wang, Haobo Zuo 외

Visual object tracking has boosted extensive intelligent applications for unmanned aerial vehicles (UAVs). However, the state-of-the-art (SOTA) enhancers for nighttime UAV tracking always neglect the uneven light distrib…

Object Trackingparameter estimationVisual Object Tracking

Taming Real-World Space-Time Video Super-Resolution with One-Step Diffusion

2026-01-28 · Shuoyan Wei, Feng Li, Chen Zhou, Runmin Cong 외 arxiv

Diffusion models have demonstrated exceptional success in video super-resolution (VSR), exhibiting powerful capabilities for generating fine-grained details. However, their potential for space-time video super-resolution…

Space-time Video Super-resolution

Tinker: Diffusion's Gift to 3D--Multi-View Consistent Editing From Sparse Inputs without Per-Scene Optimization

2025-08-20 · Canyu Zhao, Xiaoman Li, Tianjian Feng, Zhiyue Zhao 외 arxiv

We introduce Tinker, a versatile framework for high-fidelity 3D editing that operates in both one-shot and few-shot regimes without any per-scene finetuning. Unlike prior techniques that demand extensive per-scene optimi…

Generative Detail Enhancement for Physically Based Materials

2025-02-19 · Saeed Hadadan, Benedikt Bitterli, Tizian Zeltner, Jan Novák 외

We present a tool for enhancing the detail of physically based materials using an off-the-shelf diffusion model and inverse rendering. Our goal is to enhance the visual fidelity of materials with detail that is often ted…

Inverse Rendering

Diffusion-Guided Gaussian Splatting for Large-Scale Unconstrained 3D Reconstruction and Novel View Synthesis

2025-04-02 · Niluthpol Chowdhury Mithun, Tuan Pham, Qiao Wang, Ben Southall 외

Recent advancements in 3D Gaussian Splatting (3DGS) and Neural Radiance Fields (NeRF) have achieved impressive results in real-time 3D reconstruction and novel view synthesis. However, these methods struggle in large-sca…

3DGS3D ReconstructionNeRFNovel View Synthesis