Path Independent Equilibrium Models Can Better Exploit Test-Time Computation
Designing networks capable of attaining better performance with an increased inference budget is important to facilitate generalization to harder problem instances. Recent efforts have shown promising results in this direction by making use of depth-wise recurrent networks. We show that a broad class of architectures named equilibrium models display strong upwards generalization, and find that stronger performance on harder examples (which require more iterations of inference to get correct) strongly correlates with the path independence of the system -- its tendency to converge to the same steady-state behaviour regardless of initialization, given enough computation. Experimental interventions made to promote path independence result in improved generalization on harder problem instances, while those that penalize it degrade this ability. Path independence analyses are also useful on a per-example basis: for equilibrium models that have good in-distribution performance, path independence on out-of-distribution samples strongly correlates with accuracy. Our results help explain why equilibrium models are capable of strong upwards generalization and motivates future work that harnesses path independence as a general modelling principle to facilitate scalable test-time usage.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Satisficing Paths and Independent Multi-Agent Reinforcement Learning in Stochastic Games
In multi-agent reinforcement learning (MARL), independent learners are those that do not observe the actions of other agents in the system. Due to the decentralization of information, it is challenging to design independ…
Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Bootstrap Cointegration Tests in ARDL Models
The paper proposes a new bootstrap approach to the Pesaran, Shin and Smith's bound tests in a conditional equilibrium correction model with the aim to overcome some typical drawbacks of the latter, such as inconclusive i…
VISER: A Tractable Solution Concept for Games with Information Asymmetry
Many real-world games suffer from information asymmetry: one player is only aware of their own payoffs while the other player has the full game information. Examples include the critical domain of security games and adve…
Multi-agent Reinforcement LearningRobust Communication Between Parties with Nearly Independent Preferences
We study finite-state communication games in which the sender's preference is perturbed by random private idiosyncrasies. Persuasion is generically impossible within the class of statistically independent sender/receiver…
Neural Non-Equilibrium Hamiltonian Monte Carlo for Corrected Boltzmann Sampling
Sampling from an unnormalized Boltzmann density requires proposals that move probability mass globally while retaining enough path-probability information for statistical correction. We introduce Neural Non-Equilibrium H…