paper-with-me

Papers

Improving Multimodal Distillation for 3D Semantic Segmentation under Domain Shift

2025-11-21 · Björn Michele, Alexandre Boulch, Gilles Puy, Tuan-Hung Vu, Renaud Marlet, Nicolas Courty arxiv

Semantic segmentation networks trained under full supervision for one type of lidar fail to generalize to unseen lidars without intervention. To reduce the performance gap under domain shifts, a recent trend is to leverage vision foundation models (VFMs) providing robust features across domains. In this work, we conduct an exhaustive study to identify recipes for exploiting VFMs in unsupervised domain adaptation for semantic segmentation of lidar point clouds. Building upon unsupervised image-to-lidar knowledge distillation, our study reveals that: (1) the architecture of the lidar backbone is key to maximize the generalization performance on a target domain; (2) it is possible to pretrain a single backbone once and for all, and use it to address many domain shifts; (3) best results are obtained by keeping the pretrained backbone frozen and training an MLP head for semantic segmentation. The resulting pipeline achieves state-of-the-art results in four widely-recognized and challenging settings. The code will be available at: https://github.com/valeoai/muddos.

📄 PDF Abstract BibTeX arXiv:2511.17455

Code (0)

등록된 구현이 없습니다.

Tasks

Unsupervised Domain Adaptation3D Semantic SegmentationKnowledge DistillationPoint Clouds

Similar Papers 제목 키워드 기반

CM-MaskSD: Cross-Modality Masked Self-Distillation for Referring Image Segmentation

2023-05-19 · Wenxuan Wang, Jing Liu, Xingjian He, Yisi Zhang 외

Referring image segmentation (RIS) is a fundamental vision-language task that intends to segment a desired object from an image based on a given natural language expression. Due to the essentially distinct data propertie…

Image SegmentationSegmentationSemantic Segmentation

CUS3D :CLIP-based Unsupervised 3D Segmentation via Object-level Denoise

2024-09-21 · Fuyang Yu, Runze Tian, Zhen Wang, Xiaochuan Wang 외

To ease the difficulty of acquiring annotation labels in 3D data, a common method is using unsupervised and open-vocabulary semantic segmentation, which leverage 2D CLIP semantic knowledge. In this paper, unlike previous…

Open Vocabulary Semantic SegmentationOpen-Vocabulary Semantic SegmentationSegmentationSemantic Segmentation+1

Spirit Distillation: Precise Real-time Semantic Segmentation of Road Scenes with Insufficient Data

2021-03-25 · Zhiyuan Wu, Yu Jiang, Chupeng Cui, Zongmin Yang 외

Semantic segmentation of road scenes is one of the key technologies for realizing autonomous driving scene perception, and the effectiveness of deep Convolutional Neural Networks(CNNs) for this task has been demonstrated…

Autonomous DrivingFew-Shot LearningKnowledge DistillationReal-Time Semantic Segmentation+3

Exploring Generalizable Distillation for Efficient Medical Image Segmentation

2022-07-26 · Xingqun Qi, Zhuojie Wu, Min Ren, Muyi Sun 외

Efficient medical image segmentation aims to provide accurate pixel-wise predictions for medical images with a lightweight implementation framework. However, lightweight frameworks generally fail to achieve superior perf…

DecoderImage SegmentationKnowledge DistillationMedical Image Segmentation+3

Compass: Degradation-Simulated Reciprocal Learning with Lightweight Needle RWKV for Multimodal Crack Segmentation under Missing Modalities

2026-08-04 · Hui Liu, Chen Jia, Fan Shi, Xu Cheng 외 arxiv

In multimodal crack segmentation for industrial facilities, the key challenge is preventing missing modalities from degrading pixel-level performance while maintaining low computational cost. Existing methods struggle to…

Crack Segmentation