paper-with-me

Papers

Multi-Modal Multi-Task (M3T) Federated Foundation Models for Embodied AI: Potentials and Challenges for Edge Integration

2025-05-16 · Kasra Borazjani, Payam Abdisarabshali, Fardis Nadimi, Naji Khosravan, Minghui LiWang, Xianbin Wang, Yiguang Hong, Seyyedali Hosseinalipour

As embodied AI systems become increasingly multi-modal, personalized, and interactive, they must learn effectively from diverse sensory inputs, adapt continually to user preferences, and operate safely under resource and privacy constraints. These challenges expose a pressing need for machine learning models capable of swift, context-aware adaptation while balancing model generalization and personalization. Here, two methods emerge as suitable candidates, each offering parts of these capabilities: Foundation Models (FMs) provide a pathway toward generalization across tasks and modalities, whereas Federated Learning (FL) offers the infrastructure for distributed, privacy-preserving model updates and user-level model personalization. However, when used in isolation, each of these approaches falls short of meeting the complex and diverse capability requirements of real-world embodied environments. In this vision paper, we introduce Federated Foundation Models (FFMs) for embodied AI, a new paradigm that unifies the strengths of multi-modal multi-task (M3T) FMs with the privacy-preserving distributed nature of FL, enabling intelligent systems at the wireless edge. We collect critical deployment dimensions of FFMs in embodied AI ecosystems under a unified framework, which we name "EMBODY": Embodiment heterogeneity, Modality richness and imbalance, Bandwidth and compute constraints, On-device continual learning, Distributed control and autonomy, and Yielding safety, privacy, and personalization. For each, we identify concrete challenges and envision actionable research directions. We also present an evaluation framework for deploying FFMs in embodied AI systems, along with the associated trade-offs.

📄 PDF Abstract BibTeX arXiv:2505.11191

Code (0)

등록된 구현이 없습니다.

Tasks

Continual LearningFederated LearningPrivacy Preserving

Similar Papers 제목 키워드 기반

FEDKIM: Adaptive Federated Knowledge Injection into Medical Foundation Models

2024-08-17 · Xiaochen Wang, Jiaqi Wang, Houping Xiao, Jinghui Chen 외

Foundation models have demonstrated remarkable capabilities in handling diverse modalities and tasks, outperforming conventional artificial intelligence (AI) approaches that are highly task-specific and modality-reliant.…

Federated LearningMixture-of-Experts

FedLAB: Traceable Semantic Codebooks for Federated Multimodal Graph Foundation Learning

2026-06-30 · Zekai Chen, Kairui Yang, Xuaner Chen, Xunkai Li 외 arxiv

Multimodal graph foundation models aim to learn reusable knowledge from graphs enriched with text, images, attributes, and relational topology, thereby supporting diverse graph-centric and modality-centric tasks. In prac…

Representation Learning

Multi-Modal Multi-Task Federated Foundation Models for Next-Generation Extended Reality Systems: Towards Privacy-Preserving Distributed Intelligence in AR/VR/MR

2025-06-06 · Fardis Nadimi, Payam Abdisarabshali, Kasra Borazjani, Jacob Chakareski 외

Extended reality (XR) systems, which consist of virtual reality (VR), augmented reality (AR), and mixed reality (XR), offer a transformative interface for immersive, multi-modal, and embodied human-computer interaction. …

Federated LearningMixed RealityPrivacy Preserving

TAP: Two-Stage Adaptive Personalization of Multi-Task and Multi-Modal Foundation Models in Federated Learning

2025-09-30 · Seohyun Lee, Wenzhi Fang, Dong-Jun Han, Seyyedali Hosseinalipour 외 arxiv

In federated learning (FL), local personalization of models has received significant attention, yet personalized fine-tuning of foundation models remains underexplored. In particular, there is a lack of understanding in …

Federated Learning

Hierarchical Federated Foundation Models over Wireless Networks for Multi-Modal Multi-Task Intelligence: Integration of Edge Learning with D2D/P2P-Enabled Fog Learning Architectures

2025-09-03 · Payam Abdisarabshali, Fardis Nadimi, Kasra Borazjani, Naji Khosravan 외 arxiv

The rise of foundation models (FMs) has reshaped the landscape of machine learning. As these models continued to grow, leveraging geo-distributed data from wireless devices has become increasingly critical, giving rise t…