paper-with-me

홈 › Papers

Learning Robust Control Policies for Inverted Pose on Miniature Blimp Robots

2026-02-27 · Yuanlin Yang, Lin Hong, Fumin Zhang arxiv

The ability to achieve and maintain inverted poses is essential for unlocking the full agility of miniature blimp robots (MBRs). However, developing reliable inverted control strategies for MBRs remains challenging due to their complex and underactuated dynamics. To address this challenge, we propose a novel framework that enables robust control policy learning for inverted pose on MBRs. The proposed framework consists of three core stages. First, a high-fidelity three-dimensional (3D) simulation environment is constructed and calibrated using real-world MBR motion data. Second, a robust inverted control policy is trained in simulation using a modified Twin Delayed Deep Deterministic Policy Gradient (TD3) algorithm combined with a domain randomization strategy. Third, a mapping layer is designed to bridge the sim-to-real gap and facilitate real-world deployment of the learned policy. Comprehensive evaluations in the simulation environment demonstrate that the learned policy achieves a higher success rate compared to the energy-shaping controller. Furthermore, experimental results confirm that the learned policy with a mapping layer enables an MBR to achieve and maintain a fully inverted pose in real-world settings.

📄 PDF Abstract BibTeX arXiv:2602.23972

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Lightweight Tracking Control for Computationally Constrained Aerial Systems with the Newton-Raphson Method

2025-08-19 · Evanns Morales-Cuadrado, Luke Baird, Yorai Wardi, Samuel Coogan arxiv

We investigate the performance of a lightweight tracking controller, based on a flow version of the Newton-Raphson method, applied to a miniature blimp and a mid-size quadrotor. This tracking technique admits theoretical…

Learning NEAT Emergent Behaviors in Robot Swarms

2023-09-26 · Pranav Rajbhandari, Donald Sofge

When researching robot swarms, many studies observe complex group behavior emerging from the individual agents' simple local actions. However, the task of learning an individual policy to produce a desired group behavior…

MiNI-Q: A Miniature, Wire-Free Quadruped with Unbounded, Independently Actuated Leg Joints

2026-03-12 · Daniel Koh, Suraj Shah, Yufeng Wu, Dennis Hong arxiv

Physical joint limits are common in legged robots and can restrict workspace, constrain gait design, and increase the risk of hardware damage. This paper introduces MiNI-Q^2, a miniature, wire-free quadruped robot with i…

Autonomous Docking of Multi-Rotor UAVs on Blimps under the Influence of Wind Gusts

2025-11-24 · Pascal Goldschmid, Aamir Ahmad arxiv

Multi-rotor UAVs face limited flight time due to battery constraints. Autonomous docking on blimps with onboard battery recharging and data offloading offers a promising solution for extended UAV missions. However, the v…

Bi-Level Reinforcement Learning Control for an Underactuated Blimp via Center-of-Mass Reconfiguration

2026-05-02 · Xiaorui Wang, Hongwu Wang, Yue Fan, Hao Cheng 외 arxiv

This paper investigates goal-directed tracking control of underactuated blimps with center-of-mass (CoM) reconfiguration. Unlike conventional overactuated blimp designs that rely on redundant actuation for simplified con…

Reinforcement Learning