paper-with-me

Papers

MMCD: Multi-Modal Collaborative Decision-Making for Connected Autonomy with Knowledge Distillation

2025-09-19 · Rui Liu, Zikang Wang, Peng Gao, Yu Shen, Pratap Tokekar, Ming Lin arxiv

Autonomous systems have advanced significantly, but challenges persist in accident-prone environments where robust decision-making is crucial. A single vehicle's limited sensor range and obstructed views increase the likelihood of accidents. Multi-vehicle connected systems and multi-modal approaches, leveraging RGB images and LiDAR point clouds, have emerged as promising solutions. However, existing methods often assume the availability of all data modalities and connected vehicles during both training and testing, which is impractical due to potential sensor failures or missing connected vehicles. To address these challenges, we introduce a novel framework MMCD (Multi-Modal Collaborative Decision-making) for connected autonomy. Our framework fuses multi-modal observations from ego and collaborative vehicles to enhance decision-making under challenging conditions. To ensure robust performance when certain data modalities are unavailable during testing, we propose an approach based on cross-modal knowledge distillation with a teacher-student model structure. The teacher model is trained with multiple data modalities, while the student model is designed to operate effectively with reduced modalities. In experiments on $\textit{connected autonomous driving with ground vehicles}$ and $\textit{aerial-ground vehicles collaboration}$, our method improves driving safety by up to ${\it 20.7}\%$, surpassing the best-existing baseline in detecting potential accidents and making safe driving decisions. More information can be found on our website https://ruiiu.github.io/mmcd.

📄 PDF Abstract BibTeX arXiv:2509.18198

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge DistillationAutonomous DrivingPoint Clouds

Similar Papers 제목 키워드 기반

Joint Similarity Item Exploration and Overlapped User Guidance for Multi-Modal Cross-Domain Recommendation

2025-02-22 · Weiming Liu, Chaochao Chen, Jiahe Xu, Xinting Liao 외

Cross-Domain Recommendation (CDR) has been widely investigated for solving long-standing data sparsity problem via knowledge sharing across domains. In this paper, we focus on the Multi-Modal Cross-Domain Recommendation …

Collaborative FilteringDomain AdaptationTransfer Learning

Prior-guided Fusion of Multimodal Features for Change Detection from Optical-SAR Images

2026-04-07 · Xuanguang Liu, Lei Ding, Yujie Li, Chenguang Dai 외 arxiv

Multimodal change detection (MMCD) identifies changed areas in multimodal remote sensing data, demonstrating significant application value in land use monitoring and urban sustainable development. However, literature MMC…

Feature ImportanceChange Detection

Consistency-aware Fake Videos Detection on Short Video Platforms

2025-04-30 · Junxi Wang, Jize liu, Na Zhang, Yaxiong Wang

This paper focuses to detect the fake news on the short video platforms. While significant research efforts have been devoted to this task with notable progress in recent years, current detection accuracy remains subopti…

Large Language ModelMultimodal Large Language ModelPseudo Label

A Multi-Modal Contrastive Diffusion Model for Therapeutic Peptide Generation

2023-12-25 · Yongkang Wang, Xuan Liu, Feng Huang, Zhankun Xiong 외

Therapeutic peptides represent a unique class of pharmaceutical agents crucial for the treatment of human diseases. Recently, deep generative models have exhibited remarkable potential for generating therapeutic peptides…

Contrastive LearningDiversity

A Demonstration of Adaptive Collaboration of Large Language Models for Medical Decision-Making

2024-10-31 · Yubin Kim, Chanwoo Park, Hyewon Jeong, Cristina Grau-Vilchez 외

Medical Decision-Making (MDM) is a multi-faceted process that requires clinicians to assess complex multi-modal patient data patient, often collaboratively. Large Language Models (LLMs) promise to streamline this process…

Decision MakingDiagnostic