Coordinating Planning and Tracking in Layered Control Policies via Actor-Critic Learning
We propose a reinforcement learning (RL)-based algorithm to jointly train (1) a trajectory planner and (2) a tracking controller in a layered control architecture. Our algorithm arises naturally from a rewrite of the underlying optimal control problem that lends itself to an actor-critic learning approach. By explicitly learning a \textit{dual} network to coordinate the interaction between the planning and tracking layers, we demonstrate the ability to achieve an effective consensus between the two components, leading to an interpretable policy. We theoretically prove that our algorithm converges to the optimal dual network in the Linear Quadratic Regulator (LQR) setting and empirically validate its applicability to nonlinear systems through simulation experiments on a unicycle model.
Code (1)
Tasks
Reinforcement Learning (RL)Similar Papers 제목 키워드 기반
Layered Multirate Control of Constrained Linear Systems
Layered control architectures have been a standard paradigm for efficiently managing complex constrained systems. A typical architecture consists of: i) a higher layer, where a low-frequency planner controls a simple mod…
Collision AvoidanceMotion PlanningMUSIC: Learning Muscle-Driven Dexterous Hand Control
We present a data-driven approach for physics-based, muscle-driven dexterous control that enables musculoskeletal hands to perform precise piano playing for novel pieces of music outside the reference dataset. Our approa…
Reinforcement LearningMotion SynthesisFull Stack Navigation, Mapping, and Planning for the Lunar Autonomy Challenge
We present a modular, full-stack autonomy system for lunar surface navigation and mapping developed for the Lunar Autonomy Challenge. Operating in a GNSS-denied, visually challenging environment, our pipeline integrates …
Semantic SegmentationMotion PlanningVisual OdometryRisk Bounded Nonlinear Robot Motion Planning With Integrated Perception & Control
Robust autonomy stacks require tight integration of perception, motion planning, and control layers, but these layers often inadequately incorporate inherent perception and prediction uncertainties, either ignoring them …
Model Predictive ControlMotion PlanningTrack A*: Fast Visibility-Aware Trajectory Planning for Active Target Tracking
Offline reference trajectories for active target tracking are needed both for building multi-modal tracking datasets and for benchmarking online tracking planners under repeatable conditions. We present Track A star (TA …
Trajectory Planning