paper-with-me

Papers

ThermoSplat: Cross-Modal 3D Gaussian Splatting with Feature Modulation and Geometry Decoupling

2026-01-22 · Zhaoqi Su, Shihai Chen, Xinyan Lin, Liqin Huang, Zhipeng Su, Xiaoqiang Lu arxiv

Multi-modal scene reconstruction integrating RGB and thermal infrared data is essential for robust environmental perception across diverse lighting and weather conditions. However, extending 3D Gaussian Splatting (3DGS) to multi-spectral scenarios remains challenging. Current approaches often struggle to fully leverage the complementary information of multi-modal data, typically relying on mechanisms that either tend to neglect cross-modal correlations or leverage shared representations that fail to adaptively handle the complex structural correlations and physical discrepancies between spectrums. To address these limitations, we propose ThermoSplat, a novel framework that enables deep spectral-aware reconstruction through active feature modulation and adaptive geometry decoupling. First, we introduce a Spectrum-Aware Adaptive Modulation that dynamically conditions shared latent features on thermal structural priors, effectively guiding visible texture synthesis with reliable cross-modal geometric cues. Second, to accommodate modality-specific geometric inconsistencies, we propose a Modality-Adaptive Geometric Decoupling scheme that learns independent opacity offsets and executes an independent rasterization pass for the thermal branch. Additionally, a hybrid rendering pipeline is employed to integrate explicit Spherical Harmonics with implicit neural decoding, ensuring both semantic consistency and high-frequency detail preservation. Extensive experiments on the RGBT-Scenes dataset demonstrate that ThermoSplat achieves state-of-the-art rendering quality across both visible and thermal spectrums.

📄 PDF Abstract BibTeX arXiv:2601.15897

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GaussianCross: Cross-modal Self-supervised 3D Representation Learning via Gaussian Splatting

2025-08-04 · Lei Yao, Yi Wang, Yi Zhang, Moyun Liu 외 arxiv

The significance of informative and robust point representations has been widely acknowledged for 3D scene understanding. Despite existing self-supervised pre-training counterparts demonstrating promising performance, th…

Representation LearningInstance SegmentationScene UnderstandingPoint Clouds

CUS-GS: A Compact Unified Structured Gaussian Splatting Framework for Multimodal Scene Representation

2025-11-22 · Yuhang Ming, Chenxin Fang, Xingyuan Yu, Fan Zhang 외 arxiv

Recent advances in Gaussian Splatting based 3D scene representation have shown two major trends: semantics-oriented approaches that focus on high-level understanding but lack explicit 3D geometry modeling, and structure-…

GSPR: Multimodal Place Recognition Using 3D Gaussian Splatting for Autonomous Driving

2024-10-01 · Zhangshuo Qi, Junyi Ma, Jingyi Xu, Zijie Zhou 외

Place recognition is a crucial module to ensure autonomous vehicles obtain usable localization information in GPS-denied environments. In recent years, multimodal place recognition methods have gained increasing attentio…

Autonomous DrivingAutonomous Vehicles

M3: 3D-Spatial MultiModal Memory

2025-03-20 · Xueyan Zou, Yuchen Song, Ri-Zhao Qiu, Xuanbin Peng 외

We present 3D Spatial MultiModal Memory (M3), a multimodal memory system designed to retain information about medium-sized static scenes through video sources for visual perception. By integrating 3D Gaussian Splatting t…

Feature Splatting

TIGaussian: Disentangle Gaussians for Spatial-Awared Text-Image-3D Alignment

2026-01-27 · Jiarun Liu, Qifeng Chen, Yiru Zhao, Minghua Liu 외 arxiv

While visual-language models have profoundly linked features between texts and images, the incorporation of 3D modality data, such as point clouds and 3D Gaussians, further enables pretraining for 3D-related tasks, e.g.,…

Cross-Modal RetrievalScene RecognitionPoint Clouds