paper-with-me

홈 › Papers

EMind: A Foundation Model for Multi-task Electromagnetic Signals Understanding

2025-08-26 · Luqing Luo, Wenjin Gui, Yunfei Liu, Ziyue Zhang, Yunxi Zhang, Fengxiang Wang, Zonghao Guo, Zizhi Ma, Xinzhu Liu, Hanxiang He, Jinhai Li, Xin Qiu, Wupeng Xie, Yangang Sun arxiv

Deep understanding of electromagnetic signals is fundamental to dynamic spectrum management, intelligent transportation, autonomous driving and unmanned vehicle perception. The field faces challenges because electromagnetic signals differ greatly from text and images, showing high heterogeneity, strong background noise and complex joint time frequency structure, which prevents existing general models from direct use. Electromagnetic communication and sensing tasks are diverse, current methods lack cross task generalization and transfer efficiency, and the scarcity of large high quality datasets blocks the creation of a truly general multitask learning framework. To overcome these issue, we introduce EMind, an electromagnetic signals foundation model that bridges large scale pretraining and the unique nature of this modality. We build the first unified and largest standardized electromagnetic signal dataset covering multiple signal types and tasks. By exploiting the physical properties of electromagnetic signals, we devise a length adaptive multi-signal packing method and a hardware-aware training strategy that enable efficient use and representation learning from heterogeneous multi-source signals. Experiments show that EMind achieves strong performance and broad generalization across many downstream tasks, moving decisively from task specific models to a unified framework for electromagnetic intelligence. The code is available at: https://github.com/GabrielleTse/EMind.

📄 PDF Abstract BibTeX arXiv:2508.18785

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningAutonomous Driving

Similar Papers 제목 키워드 기반

WaveMind: Towards a Conversational EEG Foundation Model Aligned to Textual and Visual Modalities

2025-09-26 · Ziyi Zeng, Zhenyang Cai, Yixi Cai, Xidong Wang 외 arxiv

Electroencephalography (EEG) interpretation using multimodal large language models (MLLMs) offers a novel approach for analyzing brain signals. However, the complex nature of brain activity introduces critical challenges…

Representation Learning

PReD: An LLM-based Foundation Multimodal Model for Electromagnetic Perception, Recognition, and Decision

2026-03-30 · Zehua Han, Jing Xiao, Yiqi Duan, Mengyu Xiang 외 arxiv

Multimodal Large Language Models have demonstrated powerful cross-modal understanding and reasoning capabilities in general domains. However, in the electromagnetic (EM) domain, they still face challenges such as data sc…

The Society of HiveMind: Multi-Agent Optimization of Foundation Model Swarms to Unlock the Potential of Collective Intelligence

2025-03-07 · Noah Mamie, Susie Xi Rao

Multi-agent systems address issues of accessibility and scalability of artificial intelligence (AI) foundation models, which are often represented by large language models. We develop a framework - the "Society of HiveMi…

Logical ReasoningWorld Knowledge

TableMind++: An Uncertainty-Aware Programmatic Agent for Tool-Augmented Table Reasoning

2026-03-08 · Mingyue Cheng, Shuo Yu, Chuang Jiang, Xiaoyu Tao 외 arxiv

Table reasoning requires models to jointly perform semantic understanding and precise numerical operations. Most existing methods rely on a single-turn reasoning paradigm over tables which suffers from context overflow a…

Reinforcement Learning

StableMind: Source-Free Cross-Subject fMRI Decoding with Regularized Adaptation

2026-05-04 · Jintao Guo, Lin Wang, Shumeng Li, Jian Zhang 외 arxiv

Existing cross-subject fMRI decoding methods typically train a model on multiple scanned subjects and then adapt it to a new subject using substantial paired fMRI-image data. However, in realistic scenarios, new-subject …

Image Retrieval