paper-with-me

Papers

ELiTe: Efficient Image-to-LiDAR Knowledge Transfer for Semantic Segmentation

2024-05-07 · Zhibo Zhang, Ximing Yang, Weizhong Zhang, Cheng Jin

Cross-modal knowledge transfer enhances point cloud representation learning in LiDAR semantic segmentation. Despite its potential, the \textit{weak teacher challenge} arises due to repetitive and non-diverse car camera images and sparse, inaccurate ground truth labels. To address this, we propose the Efficient Image-to-LiDAR Knowledge Transfer (ELiTe) paradigm. ELiTe introduces Patch-to-Point Multi-Stage Knowledge Distillation, transferring comprehensive knowledge from the Vision Foundation Model (VFM), extensively trained on diverse open-world images. This enables effective knowledge transfer to a lightweight student model across modalities. ELiTe employs Parameter-Efficient Fine-Tuning to strengthen the VFM teacher and expedite large-scale model training with minimal costs. Additionally, we introduce the Segment Anything Model based Pseudo-Label Generation approach to enhance low-quality image labels, facilitating robust semantic representations. Efficient knowledge transfer in ELiTe yields state-of-the-art results on the SemanticKITTI benchmark, outperforming real-time inference models. Our approach achieves this with significantly fewer parameters, confirming its effectiveness and efficiency.

📄 PDF Abstract BibTeX arXiv:2405.04121

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge DistillationLIDAR Semantic Segmentationparameter-efficient fine-tuningPseudo LabelRepresentation LearningSemantic SegmentationTransfer Learning

Methods 이 논문이 사용한 방법론

Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…

Similar Papers 제목 키워드 기반

ELITE: Experiential Learning and Intent-Aware Transfer for Self-improving Embodied Agents

2026-03-25 · Bingqing Wei, Zhongyu Xia, Dingai Liu, Xiaoyu Zhou 외 arxiv

Vision-language models (VLMs) have shown remarkable general capabilities, yet embodied agents built on them fail at complex tasks, often skipping critical steps, proposing invalid actions, and repeating mistakes. These f…

MM-Retinal V2: Transfer an Elite Knowledge Spark into Fundus Vision-Language Pretraining

2025-01-27 · Ruiqi Wu, Na Su, Chenran Zhang, Tengfei Ma 외

Vision-language pretraining (VLP) has been investigated to generalize across diverse downstream tasks for fundus image analysis. Although recent methods showcase promising achievements, they significantly rely on large-s…

Contrastive LearningTransfer Learning

Attention-Guided Lidar Segmentation and Odometry Using Image-to-Point Cloud Saliency Transfer

2023-08-28 · Guanqun Ding, Nevrez Imamoglu, Ali Caglayan, Masahiro Murakawa 외

LiDAR odometry estimation and 3D semantic segmentation are crucial for autonomous driving, which has achieved remarkable advances recently. However, these tasks are challenging due to the imbalance of points in different…

3D Semantic SegmentationAutonomous DrivingSegmentationSemantic Segmentation+1

ProtoTransfer: Cross-Modal Prototype Transfer for Point Cloud Segmentation

2023-01-01 · ICCV 2023 1 · Pin Tang, Hai-Ming Xu, Chao Ma

Knowledge transfer from multi-modal, i.e., LiDAR points and images, to a single LiDAR modal can take advantage of complimentary information from modal-fusion but keep a single modal inference speed, showing a promisi…

Autonomous DrivingPoint Cloud SegmentationSemantic SegmentationTransfer Learning

Transfer Learning from Synthetic to Real LiDAR Point Cloud for Semantic Segmentation

2021-07-12 · Aoran Xiao, Jiaxing Huang, Dayan Guan, Fangneng Zhan 외

Knowledge transfer from synthetic to real data has been widely studied to mitigate data annotation constraints in various computer vision tasks such as semantic segmentation. However, the study focused on 2D images and i…

3D Unsupervised Domain AdaptationData AugmentationDomain AdaptationSemantic Segmentation+5