paper-with-me

홈 › Papers

Semantic-aware Network for Aerial-to-Ground Image Synthesis

2023-08-14 · Jinhyun Jang, Taeyong Song, Kwanghoon Sohn

Aerial-to-ground image synthesis is an emerging and challenging problem that aims to synthesize a ground image from an aerial image. Due to the highly different layout and object representation between the aerial and ground images, existing approaches usually fail to transfer the components of the aerial scene into the ground scene. In this paper, we propose a novel framework to explore the challenges by imposing enhanced structural alignment and semantic awareness. We introduce a novel semantic-attentive feature transformation module that allows to reconstruct the complex geographic structures by aligning the aerial feature to the ground layout. Furthermore, we propose semantic-aware loss functions by leveraging a pre-trained segmentation network. The network is enforced to synthesize realistic objects across various classes by separately calculating losses for different classes and balancing them. Extensive experiments including comparisons with previous methods and ablation studies show the effectiveness of the proposed framework both qualitatively and quantitatively.

📄 PDF Abstract BibTeX arXiv:2308.06945

Code (1)

jinhyunj/sanet 공식 구현 pytorch

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

fail 설명 없음

Similar Papers 제목 키워드 기반

Can Vision-Language Models Think from the Sky? Unifying UAV Reasoning and Generation

2026-04-07 · Jintao Sun, Gangyi Ding, Donglin Di, Hu Zhang 외 arxiv

Vision-Language Models have achieved strong progress in ground-view visual understanding, yet they remain brittle in high-altitude Unmanned Aerial Vehicle scenes, where objects are tiny and densely packed, textures are r…

Semantic Segmentation

Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis

2026-03-26 · Yancheng Zhang, Xiaohan Zhang, Guangyu Sun, Zonglin Lyu 외 arxiv

Cross-view geo-spatial learning consists of two important tasks: Cross-View Geo-Localization (CVGL) and Cross-View Image Synthesis (CVIS), both of which rely on establishing geometric correspondences between ground and a…

3D Reconstruction

SkyDiffusion: Ground-to-Aerial Image Synthesis with Diffusion Models and BEV Paradigm

2024-08-03 · Junyan Ye, Jun He, Weijia Li, Zhutao Lv 외

Ground-to-aerial image synthesis focuses on generating realistic aerial images from corresponding ground street view images while maintaining consistent content layout, simulating a top-down view. The significant viewpoi…

Image GenerationSSIM

Drone-assisted Road Gaussian Splatting with Cross-view Uncertainty

2024-08-27 · Saining Zhang, Baijun Ye, Xiaoxue Chen, Yuantao Chen 외

Robust and realistic rendering for large-scale road scenes is essential in autonomous driving simulation. Recently, 3D Gaussian Splatting (3D-GS) has made groundbreaking progress in neural rendering, but the general fide…

Autonomous DrivingNeural RenderingNovel View Synthesis

SAN: Scale-Aware Network for Semantic Segmentation of High-Resolution Aerial Images

2019-07-06 · Jingbo Lin, WeiPeng Jing, Houbing Song

High-resolution aerial images have a wide range of applications, such as military exploration, and urban planning. Semantic segmentation is a fundamental method extensively used in the analysis of high-resolution aerial …

Semantic SegmentationVocal Bursts Intensity Prediction