paper-with-me

홈 › Papers

Enabling Long-Term Cooperation in Cross-Silo Federated Learning: A Repeated Game Perspective

2021-06-22 · Ning Zhang, Qian Ma, Xu Chen

Cross-silo federated learning (FL) is a distributed learning approach where clients of the same interest train a global model cooperatively while keeping their local data private. The success of a cross-silo FL process requires active participation of many clients. Clients in cross-silo FL aim to optimize their long-term benefits by selfishly choosing their participation levels. While there has been some work on incentivizing clients to join FL, the analysis of clients' long-term selfish participation behaviors in cross-silo FL remains largely unexplored. In this paper, we analyze the selfish participation behaviors of heterogeneous clients in cross-silo FL. Specifically, we model clients' long-term selfish participation behaviors as an infinitely repeated game. For the stage game SPFL, we derive the unique Nash equilibrium (NE), and propose a distributed algorithm for each client to calculate its equilibrium participation strategy. We show that at the NE, clients fall into at most three categories: (i) free riders, (ii) a unique partial contributor (if exists), and (iii) contributors. For the long-term interactions among clients, we derive a cooperative strategy for clients which minimizes the number of free riders while increasing the amount of local data for model training. We show that enforced by a punishment strategy, such a cooperative strategy is a subgame perfect Nash equilibrium (SPNE) of the infinitely repeated game, under which some clients who are free riders at the NE of the stage game choose to be (partial) contributors. We further propose an algorithm to calculate the optimal SPNE which minimizes the number of free riders while maximizing the amount of local data for model training. Simulation results show that our derived optimal SPNE can effectively reduce the number of free riders by up to 99.3% and increase the amount of local data for model training by up to 82.3%.

📄 PDF Abstract BibTeX arXiv:2106.11814

Code (0)

등록된 구현이 없습니다.

Tasks

Federated Learning

Similar Papers 제목 키워드 기반

Rectified Schrödinger Bridge Matching for Few-Step Visual Navigation

2026-04-07 · Wuyang Luan, Junhui Li, Weiguang Zhao, Wenjian Zhang 외 arxiv

Visual navigation is a core challenge in Embodied AI, requiring autonomous agents to translate high-dimensional sensory observations into continuous, long-horizon action trajectories. While generative policies based on d…

Visual Navigation

Inter-Level Cooperation in Hierarchical Reinforcement Learning

2019-12-05 · Abdul Rahman Kreidieh, Glen Berseth, Brandon Trabucco, Samyak Parajuli 외

Hierarchies of temporally decoupled policies present a promising approach for enabling structured exploration in complex long-term planning problems. To fully achieve this approach an end-to-end training paradigm is need…

Hierarchical Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

MARL with General Utilities via Decentralized Shadow Reward Actor-Critic

2021-05-29 · Junyu Zhang, Amrit Singh Bedi, Mengdi Wang, Alec Koppel

We posit a new mechanism for cooperation in multi-agent reinforcement learning (MARL) based upon any nonlinear function of the team's long-term state-action occupancy measure, i.e., a \emph{general utility}. This subsume…

Multi-agent Reinforcement Learning

Cooperation guides evolution in a minimal model of biological evolution

2025-04-07 · Conor Houghton

A challenging simulation of evolutionary dynamics based on a three-state cellular automaton is used as a test of how cooperation can drive the evolution of complex traits. Building on the approach of Wolfram (2025), the …

Diversity

On the Emergence of Cooperation in the Repeated Prisoner's Dilemma

2022-11-24 · Maximilian Schaefer

Using simulations between pairs of $\epsilon$-greedy q-learners with one-period memory, this article demonstrates that the potential function of the stochastic replicator dynamics (Foster and Young, 1990) allows it to pr…