paper-with-me

홈 › Papers

Dimensionless Policies based on the Buckingham $π$ Theorem: Is This a Good Way to Generalize Numerical Results?

2023-07-29 · Alexandre Girard

The answer to the question posed in the title is yes if the context (the list of variables defining the motion control problem) is dimensionally similar. This article explores the use of the Buckingham $\pi$ theorem as a tool to encode the control policies of physical systems into a more generic form of knowledge that can be reused in various situations. This approach can be interpreted as enforcing invariance to the scaling of the fundamental units in an algorithm learning a control policy. First, we show, by restating the solution to a motion control problem using dimensionless variables, that (1) the policy mapping involves a reduced number of parameters and (2) control policies generated numerically for a specific system can be transferred exactly to a subset of dimensionally similar systems by scaling the input and output variables appropriately. Those two generic theoretical results are then demonstrated, with numerically generated optimal controllers, for the classic motion control problem of swinging up a torque-limited inverted pendulum and positioning a vehicle in slippery conditions. We also discuss the concept of regime, a region in the space of context variables, that can help to relax the similarity condition. Furthermore, we discuss how applying dimensional scaling of the input and output of a context-specific black-box policy is equivalent to substituting new system parameters in an analytical equation under some conditions, using a linear quadratic regulator (LQR) and a computed torque controller as examples. It remains to be seen how practical this approach can be to generalize policies for more complex high-dimensional problems, but the early results show that it is a promising transfer learning tool for numerical approaches like dynamic programming and reinforcement learning.

📄 PDF Abstract BibTeX arXiv:2307.15852

Code (1)

alx87grd/dimensionlesspolicies 공식 구현

Tasks

Transfer Learning

Similar Papers 제목 키워드 기반

Improving Controller Generalization with Dimensionless Markov Decision Processes

2025-04-14 · Valentin Charvet, Sebastian Stein, Roderick Murray-Smith

Controllers trained with Reinforcement Learning tend to be very specialized and thus generalize poorly when their testing environment differs from their training one. We propose a Model-Based approach to increase general…

Zero-Shot Policy Transfer in Reinforcement Learning using Buckingham's Pi Theorem

2025-10-09 · Francisco Pascoa, Ian Lalonde, Alexandre Girard arxiv

Reinforcement learning (RL) policies often fail to generalize to new robots, tasks, or environments with different physical parameters, a challenge that limits their real-world applicability. This paper presents a simple…

Reinforcement Learning

Dimensionally Consistent Learning with Buckingham Pi

2022-02-09 · Joseph Bakarji, Jared Callaham, Steven L. Brunton, J. Nathan Kutz

In the absence of governing equations, dimensional analysis is a robust technique for extracting insights and finding symmetries in physical systems. Given measurement variables and parameters, the Buckingham Pi theorem …

The Algebra of Units: From Buckingham's Pi-grec Theorem to Latent-Variable Learning

2026-06-15 · Mauro Valorani arxiv

Engineers often measure many quantities-speed, pressure, temperature, length-expressed in different physical units. The Buckingham Pi-grec theorem states that these variables can always be combined into a smaller set of …

Pi theorem formulation of flood mapping

2022-11-01 · Mark S. Bartlett, Jared Van Blitterswyk, Martha Farella, Jinshu Li 외

Rapid delineation of flash flood extents is critical to mobilize emergency resources and to manage evacuations, thereby saving lives and property. Machine learning (ML) approaches enable rapid flood delineation with redu…

Managementregression