paper-with-me

홈 › Papers

Generative Autoregressive Transformers for Model-Agnostic Federated MRI Reconstruction

2025-02-06 · Valiyeh A. Nezhad, Gokberk Elmas, Bilal Kabas, Fuat Arslan, Tolga Çukur

Although learning-based models hold great promise for MRI reconstruction, single-site models built on limited local datasets often suffer from poor generalization. This challenge has spurred interest in collaborative model training on multi-site datasets via federated learning (FL) -- a privacy-preserving framework that aggregates model updates instead of sharing imaging data. Conventional FL aggregates locally trained model weights into a global model, inherently constraining all sites to use a homogeneous model architecture. This rigidity forces sites to compromise on architectures tailored to their compute resources and application-specific needs, making conventional FL unsuitable for model-heterogeneous settings where each site may prefer a distinct architecture. To overcome this limitation, we introduce FedGAT, a novel model-agnostic FL technique based on generative autoregressive transformers. FedGAT decentralizes the training of a global generative prior that learns the distribution of multi-site MR images. For high-fidelity synthesis, we propose a novel site-prompted GAT prior that controllably synthesizes realistic MR images from desired sites via autoregressive prediction across spatial scales. Each site then trains its own reconstruction model -- using an architecture of its choice -- on a hybrid dataset augmenting its local MRI dataset with GAT-generated synthetic MR images emulating datasets from other sites. This hybrid training strategy enables site-specific reconstruction models to generalize more effectively across diverse data distributions while preserving data privacy. Comprehensive experiments on multi-institutional datasets demonstrate that FedGAT enables flexible, model-heterogeneous collaborations and achieves superior within-site and cross-site reconstruction performance compared to state-of-the-art FL baselines.

📄 PDF Abstract BibTeX arXiv:2502.04521

Code (1)

icon-lab/FedGAT 공식 구현 pytorch

Tasks

Federated LearningMRI ReconstructionPrivacy Preserving

Methods 이 논문이 사용한 방법론

GAT A Graph Attention Network (GAT) is a neural network architecture that operates on graph-structured data, leveraging masked self-attentional layers to address the shortcomings…

Similar Papers 제목 키워드 기반

Autoregressive model path dependence near Ising criticality

2024-08-28 · Yi Hong Teoh, Roger G. Melko

Autoregressive models are a class of generative model that probabilistically predict the next output of a sequence based on previous inputs. The autoregressive sequence is by definition one-dimensional (1D), which is nat…

model

Trie-Aware Transformers for Generative Recommendation

2026-02-25 · Zhenxiang Xu, Jiawei Chen, Sirui Chen, Yong He 외 arxiv

Generative recommendation (GR) aligns with advances in generative AI by casting next-item prediction as token-level generation rather than score-based ranking. Most GR methods adopt a two-stage pipeline: (i) \textit{item…

Generative AI for Video Trailer Synthesis: From Extractive Heuristics to Autoregressive Creativity

2026-04-03 · Abhishek Dharmaratnakar, Srivaths Ranganathan, Debanshu Das, Anushree Sinha arxiv

The domain of automatic video trailer generation is currently undergoing a profound paradigm shift, transitioning from heuristic-based extraction methods to deep generative synthesis. While early methodologies relied hea…

Feature Engineering

Visual Prompt Tuning for Generative Transfer Learning

2022-10-03 · CVPR 2023 1 · Kihyuk Sohn, Yuan Hao, José Lezama, Luisa Polania 외

Transferring knowledge from an image synthesis model trained on a large dataset is a promising direction for learning generative image models from various domains efficiently. While previous works have studied GAN models…

Image GenerationTransfer LearningVisual Prompt Tuning

Quantization-Free Autoregressive Action Transformer

2025-03-18 · Ziyad Sheebaelhamd, Michael Tschannen, Michael Muehlebach, Claire Vernade

Current transformer-based imitation learning approaches introduce discrete action representations and train an autoregressive transformer decoder on the resulting latent code. However, the initial quantization breaks the…

Imitation LearningQuantizationSequential Decision Making