paper-with-me

Papers

Federated Cross-Modal Style-Aware Prompt Generation

2025-08-17 · Suraj Prasad, Navyansh Mahla, Sunny Gupta, Amit Sethi arxiv

Prompt learning has propelled vision-language models like CLIP to excel in diverse tasks, making them ideal for federated learning due to computational efficiency. However, conventional approaches that rely solely on final-layer features miss out on rich multi-scale visual cues and domain-specific style variations in decentralized client data. To bridge this gap, we introduce FedCSAP (Federated Cross-Modal Style-Aware Prompt Generation). Our framework harnesses low, mid, and high-level features from CLIP's vision encoder alongside client-specific style indicators derived from batch-level statistics. By merging intricate visual details with textual context, FedCSAP produces robust, context-aware prompt tokens that are both distinct and non-redundant, thereby boosting generalization across seen and unseen classes. Operating within a federated learning paradigm, our approach ensures data privacy through local training and global aggregation, adeptly handling non-IID class distributions and diverse domain-specific styles. Comprehensive experiments on multiple image classification datasets confirm that FedCSAP outperforms existing federated prompt learning methods in both accuracy and overall generalization.

📄 PDF Abstract BibTeX arXiv:2508.12399

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyImage ClassificationFederated Learning

Similar Papers 제목 키워드 기반

BadPromptFL: A Novel Backdoor Threat to Prompt-based Federated Learning in Multimodal Models

2025-08-11 · Maozhen Zhang, Mengnan Zhao, Wei Wang, Bo Wang arxiv

Prompt-based tuning has emerged as a lightweight alternative to full fine-tuning in large vision-language models, enabling efficient adaptation via learned contextual prompts. This paradigm has recently been extended to …

Federated Learning

Toward Federated Multimodal Graph Foundation Models: A Topology-Aware Multimodal Alignment Framework

2026-07-17 · Xunkai Li, Guohao Fu, Yuming Ai, Zhengyu Wu 외 arxiv

Multimodal-attributed graphs (MAGs), whose nodes carry modalities such as images and text alongside topological structure, now pervade applications including social platforms, e-commerce, and biomedical networks, offerin…

Federated LearningFew-Shot LearningGraph Learning

PRISM: Topology-Aware Cross-Modal Imputation for Modality-Deficient Federated Graph Learning

2026-06-08 · Zekai Chen, Miao Zhang, Jiayang Xing, Xunkai Li 외 arxiv

Multimodal federated graph learning (MM-FGL) aims to collaboratively learn from decentralized graphs with text and images. However, real-world clients may not share a common modality basis: a visual-search client may con…

Graph Learning

Mitigating Group-Level Fairness Disparities in Federated Visual Language Models

2025-05-03 · Chaomeng Chen, Zitong Yu, Junhao Dong, Sen Su 외

Visual language models (VLMs) have shown remarkable capabilities in multimodal tasks but face challenges in maintaining fairness across demographic groups, particularly when deployed in federated learning (FL) environmen…

counterfactualFairnessFederated LearningPrivacy Preserving

FedTaste: Topology-Aware Structural Transfer for Multimodal Federated Learning with Missing Modalities

2026-07-25 · Haochen Liang, Jie Zhang, Hideya Ochiai arxiv

Multimodal Federated Learning is often challenged by arbitrary modality missingness and Non-IID data distributions, which lead to severe representation drift and hinder effective collaboration across clients. Existing me…

Federated Learning