paper-with-me

홈 › Papers

Large Model for Small Data: Foundation Model for Cross-Modal RF Human Activity Recognition

2024-10-13 · Yuxuan Weng, Guoquan Wu, Tianyue Zheng, Yanbing Yang, Jun Luo

Radio-Frequency (RF)-based Human Activity Recognition (HAR) rises as a promising solution for applications unamenable to techniques requiring computer visions. However, the scarcity of labeled RF data due to their non-interpretable nature poses a significant obstacle. Thanks to the recent breakthrough of foundation models (FMs), extracting deep semantic insights from unlabeled visual data become viable, yet these vision-based FMs fall short when applied to small RF datasets. To bridge this gap, we introduce FM-Fi, an innovative cross-modal framework engineered to translate the knowledge of vision-based FMs for enhancing RF-based HAR systems. FM-Fi involves a novel cross-modal contrastive knowledge distillation mechanism, enabling an RF encoder to inherit the interpretative power of FMs for achieving zero-shot learning. It also employs the intrinsic capabilities of FM and RF to remove extraneous features for better alignment between the two modalities. The framework is further refined through metric-based few-shot learning techniques, aiming to boost the performance for predefined HAR tasks. Comprehensive evaluations evidently indicate that FM-Fi rivals the effectiveness of vision-based methodologies, and the evaluation results provide empirical validation of FM-Fi's generalizability across various environments.

📄 PDF Abstract BibTeX arXiv:2410.19766

Code (0)

등록된 구현이 없습니다.

Tasks

Activity RecognitionFew-Shot LearningHuman Activity RecognitionKnowledge DistillationmodelZero-Shot Learning

Methods 이 논문이 사용한 방법론

Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…

Similar Papers 제목 키워드 기반

MedMIX: Modality-Internal Expert Fusion for Multimodal Medical Diagnosis

2026-05-15 · Seungik Cho, Anqi Li, Wei Qiu arxiv

Multimodal clinical prediction faces three challenges: multiple foundation models (FMs) with complementary strengths per modality, pervasive missing modalities at training and test time, and sample-specific variation in …

Medical Diagnosis

CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models

2025-03-09 · Wei Dai, Peilin Chen, Malinda Lu, Daniel Li 외

Recent advances in clinical AI have enabled remarkable progress across many clinical domains. However, existing benchmarks and models are primarily limited to a small set of modalities and tasks, which hinders the develo…

FLORO: A Multimodal Geospatial Foundation Model for Ecological Remote Sensing Across Sensors and Scales

2026-05-27 · Jorge L. Rodriguez, Victor Angulo Morales, Areej Alwahas, Mariana Elias Lara 외 arxiv

Foundation models offer a promising route to transferable remote sensing representations, but many current approaches depend on very large pretraining datasets and fixed sensor configurations, limiting their suitability …

Scene Classification

Reverso: Efficient Time Series Foundation Models for Zero-shot Forecasting

2026-02-19 · Xinghong Fu, Yanhong Li, Georgios Papaioannou, Yoon Kim arxiv

Learning time series foundation models has been shown to be a promising approach for zero-shot time series forecasting across diverse time series domains. Insofar as scaling has been a critical driver of performance of f…

Time Series ForecastingData Augmentation

BioBridge: Bridging Biomedical Foundation Models via Knowledge Graphs

2023-10-05 · Zifeng Wang, Zichen Wang, Balasubramaniam Srinivasan, Vassilis N. Ioannidis 외

Foundation models (FMs) are able to leverage large volumes of unlabeled data to demonstrate superior performance across a wide range of tasks. However, FMs developed for biomedical domains have largely remained unimodal,…

Cross-Modal RetrievalDomain GeneralizationKnowledge GraphsQuestion Answering+1