paper-with-me

홈 › Papers

Neural Policy Composition from Free Energy Minimization

2025-12-04 · Francesca Rossi, Veronica Centorrino, Francesco Bullo, Giovanni Russo arxiv

The ability to flexibly compose previously acquired skills to execute intelligent behaviors is a hallmark of natural intelligence. Such compositional flexibility is often attributed to context-dependent gating mechanisms that determine how multiple policies or behavioral primitives are combined. Yet, despite remarkable efforts, the normative objective from which such gating rules should arise, and the neural computations capable of implementing them, remain unclear. Existing approaches typically rely on prespecified design choices for the gating rules, and remain tied to specific architectures, learning paradigms, or datasets. Here, we introduce a normative framework in which policy composition emerges from the minimization of a variational free energy, providing a principled and broadly applicable objective for gating. Based on this framework, we derive a continuous-time gradient flow whose trajectories are guaranteed to converge, with explicit rate, to the optimal composition of primitives. We further show that this dynamics admits a mechanistic neural implementation as a soft-competitive recurrent circuit with context-sensitive local interactions. We evaluate the model on emerging flocking behaviors in multi-agent systems, human decision-making in bandit tasks, and control benchmarks in layered architectures. Across these settings, the model provides interpretable mechanistic accounts of policy composition, reproduces key behavioral signatures, yields insights into data, and matches or outperforms established models.

📄 PDF Abstract BibTeX arXiv:2512.04745

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Reinforcement Learning with Subspaces using Free Energy Paradigm

2020-12-13 · Milad Ghorbani, Reshad Hosseini, Seyed Pooya Shariatpanahi, Majid Nili Ahmadabadi

In large-scale problems, standard reinforcement learning algorithms suffer from slow learning speed. In this paper, we follow the framework of using subspaces to tackle this problem. We propose a free-energy minimization…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Thompson Sampling

Expected Free Energy-based Planning as Variational Inference

2026-06-09 · Wouter W. L. Nuijten, Thijs van de Laar, Bert de Vries arxiv

Planning under uncertainty requires agents to balance goal achievement with information gathering. Active inference addresses this through the Expected Free Energy (EFE), a cost function that unifies instrumental and epi…

A Message Passing Realization of Expected Free Energy Minimization

2025-08-04 · Wouter W. L. Nuijten, Mykola Lukashchuk, Thijs van de Laar, Bert de Vries arxiv

We present a message passing approach to Expected Free Energy (EFE) minimization on factor graphs, based on the theory introduced in arXiv:2504.14898. By reformulating EFE minimization as Variational Free Energy minimiza…

Non-conflicting Energy Minimization in Reinforcement Learning based Robot Control

2025-09-01 · Skand Peri, Akhil Perincherry, Bikram Pandit, Stefan Lee arxiv

Efficient robot control often requires balancing task performance with energy expenditure. A common approach in reinforcement learning (RL) is to penalize energy use directly as part of the reward function. This requires…

Reinforcement Learning

Whence the Expected Free Energy?

2020-04-17 · Beren Millidge, Alexander Tschantz, Christopher L. Buckley

The Expected Free Energy (EFE) is a central quantity in the theory of active inference. It is the quantity that all active inference agents are mandated to minimize through action, and its decomposition into extrinsic an…