paper-with-me

Papers

Geometry-guided Cross-view Diffusion for One-to-many Cross-view Image Synthesis

2024-12-04 · Tao Jun Lin, Wenqing Wang, Yujiao Shi, Akhil Perincherry, Ankit Vora, Hongdong Li

This paper presents a novel approach for cross-view synthesis aimed at generating plausible ground-level images from corresponding satellite imagery or vice versa. We refer to these tasks as satellite-to-ground (Sat2Grd) and ground-to-satellite (Grd2Sat) synthesis, respectively. Unlike previous works that typically focus on one-to-one generation, producing a single output image from a single input image, our approach acknowledges the inherent one-to-many nature of the problem. This recognition stems from the challenges posed by differences in illumination, weather conditions, and occlusions between the two views. To effectively model this uncertainty, we leverage recent advancements in diffusion models. Specifically, we exploit random Gaussian noise to represent the diverse possibilities learnt from the target view data. We introduce a Geometry-guided Cross-view Condition (GCC) strategy to establish explicit geometric correspondences between satellite and street-view features. This enables us to resolve the geometry ambiguity introduced by camera pose between image pairs, boosting the performance of cross-view image synthesis. Through extensive quantitative and qualitative analyses on three benchmark cross-view datasets, we demonstrate the superiority of our proposed geometry-guided cross-view condition over baseline methods, including recent state-of-the-art approaches in cross-view image synthesis. Our method generates images of higher quality, fidelity, and diversity than other state-of-the-art approaches.

📄 PDF Abstract BibTeX arXiv:2412.03315

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Focus 설명 없음

Similar Papers 제목 키워드 기반

GeoQuery: Geometry-Query Diffusion for Sparse-View Reconstruction

2026-05-12 · Xiao Cao, Yuze Li, Youmin Zhang, Jiayu Song 외 arxiv

3D Gaussian Splatting (3DGS) has emerged as a prominent paradigm for 3D reconstruction and novel view synthesis. However, it remains vulnerable to severe artifacts when trained under sparse-view constraints. While recent…

Novel View Synthesis3D Reconstruction

GeoFace: Consistent Multi-View Face Generation with Geometry-Constrained Diffusion

2026-06-26 · Yeji Choi, Jinhyeok Choi, Jaewon Min, Minkyung Kwon 외 arxiv

We present GeoFace, a geometry-constrained multi-view diffusion framework for consistent face generation from a single input. % While recent multi-view diffusion models achieve photorealistic synthesis at the per-view le…

3D Reconstruction

Template-Free Single-View 3D Human Digitalization with Diffusion-Guided LRM

2024-01-22 · Zhenzhen Weng, Jingyuan Liu, Hao Tan, Zhan Xu 외

Reconstructing 3D humans from a single image has been extensively investigated. However, existing approaches often fall short on capturing fine geometry and appearance details, hallucinating occluded parts with plausible…

DecoderNeRF

DreamVTON: Customizing 3D Virtual Try-on with Personalized Diffusion Models

2024-07-23 · Zhenyu Xie, Haoye Dong, Yufei Gao, Zehua Ma 외

Image-based 3D Virtual Try-ON (VTON) aims to sculpt the 3D human according to person and clothes images, which is data-efficient (i.e., getting rid of expensive 3D data) but challenging. Recent text-to-3D methods achieve…

Text to 3DVirtual Try-on

ProDiG: Progressive Diffusion-Guided Gaussian Splatting for Aerial to Ground Reconstruction

2026-04-02 · Sirshapan Mitra, Yogesh S. Rawat arxiv

Generating ground-level views and coherent 3D site models from aerial-only imagery is challenging due to extreme viewpoint changes, missing intermediate observations, and large scale variations. Existing methods either r…