paper-with-me

Papers

Multi-Modal Self-Supervised Semantic Communication

2025-03-18 · Hang Zhao, Hongru Li, Dongfang Xu, Shenghui Song, Khaled B. Letaief

Semantic communication is emerging as a promising paradigm that focuses on the extraction and transmission of semantic meanings using deep learning techniques. While current research primarily addresses the reduction of semantic communication overhead, it often overlooks the training phase, which can incur significant communication costs in dynamic wireless environments. To address this challenge, we propose a multi-modal semantic communication system that leverages multi-modal self-supervised learning to enhance task-agnostic feature extraction. The proposed approach employs self-supervised learning during the pre-training phase to extract task-agnostic semantic features, followed by supervised fine-tuning for downstream tasks. This dual-phase strategy effectively captures both modality-invariant and modality-specific features while minimizing training-related communication overhead. Experimental results on the NYU Depth V2 dataset demonstrate that the proposed method significantly reduces training-related communication overhead while maintaining or exceeding the performance of existing supervised learning approaches. The findings underscore the advantages of multi-modal self-supervised learning in semantic communication, paving the way for more efficient and scalable edge inference systems.

📄 PDF Abstract BibTeX arXiv:2503.13940

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised LearningSemantic Communication

Similar Papers 제목 키워드 기반

Communication-Efficient Multi-Modal Edge Inference via Uncertainty-Aware Distributed Learning

2026-01-21 · Hang Zhao, Hongru Li, Dongfang Xu, Shenghui Song 외 arxiv

Semantic communication is emerging as a key enabler for distributed edge intelligence due to its capability to convey task-relevant meaning. However, achieving communication-efficient training and robust inference over w…

Self-Supervised LearningSemantic CommunicationScene Classification

Multi-Task Multi-Modal Self-Supervised Learning for Facial Expression Recognition

2024-04-16 · Marah Halawa, Florian Blume, Pia Bideau, Martin Maier 외

Human communication is multi-modal; e.g., face-to-face interaction involves auditory signals (speech) and visual signals (face movements and hand gestures). Hence, it is essential to exploit multiple modalities when desi…

Emotion ClassificationEmotion Recognition in ConversationFacial Expression RecognitionSelf-Supervised Learning

Multi-Modal Fusion-Based Multi-Task Semantic Communication System

2024-07-01 · Zengle Zhu, Rongqing Zhang, Xiang Cheng, Liuqing Yang

In recent years, there has been significant progress in semantic communication systems empowered by deep learning techniques. It has greatly improved the efficiency of information transmission. Nevertheless, traditional …

Semantic Communication

Multi-Modal Semantic Communication

2025-12-17 · Matin Mortaheb, Erciyes Karakaya, Sennur Ulukus arxiv

Semantic communication aims to transmit information most relevant to a task rather than raw data, offering significant gains in communication efficiency for applications such as telepresence, augmented reality, and remot…

Information ExtractionSemantic Communication

Self-Supervised Adversarial Hashing Networks for Cross-Modal Retrieval

2018-04-04 · CVPR 2018 6 · Chao Li, Cheng Deng, Ning li, Wei Liu 외

Thanks to the success of deep learning, cross-modal retrieval has made significant progress recently. However, there still remains a crucial bottleneck: how to bridge the modality gap to further enhance the retrieval acc…

Cross-Modal RetrievalRetrieval