Energy-Based Learning for Cooperative Games, with Applications to Valuation Problems in Machine Learning
Valuation problems, such as feature interpretation, data valuation and model valuation for ensembles, become increasingly more important in many machine learning applications. Such problems are commonly solved by well-known game-theoretic criteria, such as Shapley value or Banzhaf value. In this work, we present a novel energy-based treatment for cooperative games, with a theoretical justification by the maximum entropy framework. Surprisingly, by conducting variational inference of the energy-based model, we recover various game-theoretic valuation criteria through conducting one-step fixed point iteration for maximizing the mean-field ELBO objective. This observation also verifies the rationality of existing criteria, as they are all attempting to decouple the correlations among the players through the mean-field approach. By running fixed point iteration for multiple steps, we achieve a trajectory of the valuations, among which we define the valuation with the best conceivable decoupling error as the Variational Index. We prove that under uniform initializations, these variational valuations all satisfy a set of game-theoretic axioms. We experimentally demonstrate that the proposed Variational Index enjoys lower decoupling error and better valuation performance on certain synthetic and real-world valuation problems.
Code (0)
등록된 구현이 없습니다.
Tasks
Data ValuationVariational InferenceMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Learning in Mean Field Games: A Survey
Non-cooperative and cooperative games with a very large number of players have many applications but remain generally intractable when the number of players increases. Introduced by Lasry and Lions, and Huang, Caines and…
Reinforcement Learning (RL)SurveySCC-rFMQ Learning in Cooperative Markov Games with Continuous Actions
Although many reinforcement learning methods have been proposed for learning the optimal solutions in single-agent continuous-action domains, multiagent coordination domains with continuous actions have received relative…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Online coalitional games for real-time payoff distribution with applications to energy markets
Motivated by the markets operating on fast time scales, we present a framework for online coalitional games with time-varying coalitional values and propose real-time payoff distribution mechanisms. Specifically, we desi…
Cooperative Energy Scheduling for Microgrids under Peak Demand Energy Plans
A cooperative energy scheduling method is proposed that allows joint energy optimization for a group of microgrids to achieve cost savings that the microgrids could not achieve individually. The discussed microgrids may …
FairnessSchedulingImplementations of Cooperative Games Under Non-Cooperative Solution Concepts
Cooperative games can be distinguished as non-cooperative games in which players can freely sign binding agreements to form coalitions. These coalitions inherit a joint strategy set and seek to maximize collective payoff…
AllForm