ELiTe: Efficient Image-to-LiDAR Knowledge Transfer for Semantic Segmentation
Cross-modal knowledge transfer enhances point cloud representation learning in LiDAR semantic segmentation. Despite its potential, the \textit{weak teacher challenge} arises due to repetitive and non-diverse car camera images and sparse, inaccurate ground truth labels. To address this, we propose the Efficient Image-to-LiDAR Knowledge Transfer (ELiTe) paradigm. ELiTe introduces Patch-to-Point Multi-Stage Knowledge Distillation, transferring comprehensive knowledge from the Vision Foundation Model (VFM), extensively trained on diverse open-world images. This enables effective knowledge transfer to a lightweight student model across modalities. ELiTe employs Parameter-Efficient Fine-Tuning to strengthen the VFM teacher and expedite large-scale model training with minimal costs. Additionally, we introduce the Segment Anything Model based Pseudo-Label Generation approach to enhance low-quality image labels, facilitating robust semantic representations. Efficient knowledge transfer in ELiTe yields state-of-the-art results on the SemanticKITTI benchmark, outperforming real-time inference models. Our approach achieves this with significantly fewer parameters, confirming its effectiveness and efficiency.
Code (0)
등록된 구현이 없습니다.
Tasks
Knowledge DistillationLIDAR Semantic Segmentationparameter-efficient fine-tuningPseudo LabelRepresentation LearningSemantic SegmentationTransfer LearningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
ELITE: Experiential Learning and Intent-Aware Transfer for Self-improving Embodied Agents
Vision-language models (VLMs) have shown remarkable general capabilities, yet embodied agents built on them fail at complex tasks, often skipping critical steps, proposing invalid actions, and repeating mistakes. These f…
MM-Retinal V2: Transfer an Elite Knowledge Spark into Fundus Vision-Language Pretraining
Vision-language pretraining (VLP) has been investigated to generalize across diverse downstream tasks for fundus image analysis. Although recent methods showcase promising achievements, they significantly rely on large-s…
Contrastive LearningTransfer LearningAttention-Guided Lidar Segmentation and Odometry Using Image-to-Point Cloud Saliency Transfer
LiDAR odometry estimation and 3D semantic segmentation are crucial for autonomous driving, which has achieved remarkable advances recently. However, these tasks are challenging due to the imbalance of points in different…
3D Semantic SegmentationAutonomous DrivingSegmentationSemantic Segmentation+1ProtoTransfer: Cross-Modal Prototype Transfer for Point Cloud Segmentation
Knowledge transfer from multi-modal, i.e., LiDAR points and images, to a single LiDAR modal can take advantage of complimentary information from modal-fusion but keep a single modal inference speed, showing a promisi…
Autonomous DrivingPoint Cloud SegmentationSemantic SegmentationTransfer LearningTransfer Learning from Synthetic to Real LiDAR Point Cloud for Semantic Segmentation
Knowledge transfer from synthetic to real data has been widely studied to mitigate data annotation constraints in various computer vision tasks such as semantic segmentation. However, the study focused on 2D images and i…
3D Unsupervised Domain AdaptationData AugmentationDomain AdaptationSemantic Segmentation+5