paper-with-me

홈 › Papers

DF4LCZ: A SAM-Empowered Data Fusion Framework for Scene-Level Local Climate Zone Classification

2024-03-14 · Qianqian Wu, Xianping Ma, Jialu Sui, Man-on Pun

Recent advancements in remote sensing (RS) technologies have shown their potential in accurately classifying local climate zones (LCZs). However, traditional scene-level methods using convolutional neural networks (CNNs) often struggle to integrate prior knowledge of ground objects effectively. Moreover, commonly utilized data sources like Sentinel-2 encounter difficulties in capturing detailed ground object information. To tackle these challenges, we propose a data fusion method that integrates ground object priors extracted from high-resolution Google imagery with Sentinel-2 multispectral imagery. The proposed method introduces a novel Dual-stream Fusion framework for LCZ classification (DF4LCZ), integrating instance-based location features from Google imagery with the scene-level spatial-spectral features extracted from Sentinel-2 imagery. The framework incorporates a Graph Convolutional Network (GCN) module empowered by the Segment Anything Model (SAM) to enhance feature extraction from Google imagery. Simultaneously, the framework employs a 3D-CNN architecture to learn the spectral-spatial features of Sentinel-2 imagery. Experiments are conducted on a multi-source remote sensing image dataset specifically designed for LCZ classification, validating the effectiveness of the proposed DF4LCZ. The related code and dataset are available at https://github.com/ctrlovefly/DF4LCZ.

📄 PDF Abstract BibTeX arXiv:2403.09367

Code (1)

ctrlovefly/df4lcz 공식 구현 tf

Similar Papers 제목 키워드 기반

O$^2$-Recon: Completing 3D Reconstruction of Occluded Objects in the Scene with a Pre-trained 2D Diffusion Model

2023-08-18 · Yubin Hu, Sheng Ye, Wang Zhao, Matthieu Lin 외

Occlusion is a common issue in 3D reconstruction from RGB-D videos, often blocking the complete reconstruction of objects and presenting an ongoing problem. In this paper, we propose a novel framework, empowered by a 2D …

3D ReconstructionBlocking

LASA: Instance Reconstruction from Real Scans using A Large-scale Aligned Shape Annotation Dataset

2023-12-19 · CVPR 2024 1 · Haolin Liu, Chongjie Ye, Yinyu Nie, Yingfan He 외

Instance shape reconstruction from a 3D scene involves recovering the full geometries of multiple objects at the semantic instance level. Many methods leverage data-driven learning due to the intricacies of scene complex…

3D Object DetectionObjectobject-detectionObject Detection

Attention-based Multi-modal Fusion Network for Semantic Scene Completion

2020-03-31 · Siqi Li, Changqing Zou, Yipeng Li, Xibin Zhao 외

This paper presents an end-to-end 3D convolutional network named attention-based multi-modal fusion network (AMFNet) for the semantic scene completion (SSC) task of inferring the occupancy and semantic labels of a volume…

2D Semantic Segmentation3D Semantic Scene CompletionSegmentationSemantic Segmentation

Text-to-Image GAN with Pretrained Representations

2024-12-30 · Xiaozhou You, Jian Zhang

Generating desired images conditioned on given text descriptions has received lots of attention. Recently, diffusion models and autoregressive models have demonstrated their outstanding expressivity and gradually replace…

Domain GeneralizationImage GenerationScene Understanding

Text-Driven Fusion for Infrared and Visible Images: Achieving Image Scene Adaptation on Hyperbolic Space

2026-06-13 · Huan Kang, Hui Li, Tianyang Xu, Tao Zhou 외 arxiv

Infrared and visible image fusion aims to integrate complementary modalities, while existing Euclidean methods impose rigid distance metrics that distort multi-modal interactions and parent-to-child semantic hierarchies.…