paper-with-me

Hierarchical Reinforcement Learning

1개 벤치마크 · 논문 481편 · 이 태스크의 논문 보기 →

Benchmarks

Ant + Maze

결과 1개

Most implemented

Papers

Learning Highly Dynamic Skills Transition for Quadruped Jumping Through Constrained Space

2026-08-20 · Zeren Luo, Jiahui Zhang, Yimin Han, Ji Ma 외 arxiv

Although legged animals are capable of performing explosive motions while traversing confined spaces, replicating this behavior in quadrupedal robots has been a longstanding challenge. Here, we propose a hierarchical rei…

Hierarchical Reinforcement Learning

Enhancing Social Intelligence in LLMs with Hierarchical Reasoning and Utterance-Level Goal Rewarding

2026-08-06 · Xiaofeng Wang, Kakam Chong, Shuai Xiao, DeXin Kong 외 arxiv

Large language models (LLMs) excel in structured tasks but struggle with dynamic social interactions, where success requires long-term goal coordination and rapid adaptation. Current methods often apply uniform goal-base…

Hierarchical Reinforcement Learning

Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning

2026-07-26 · Zahra Abdalla Elashaal, Afef Hfaiedh, Nahla Khraief, Issmail Ellabib 외 arxiv

Exploration in sparse-reward long-horizon tasks poses significant challenges for reinforcement learning. To address these challenges, we propose a two-level Hierarchical Reinforcement Learning (HRL) framework. The first …

Hierarchical Reinforcement LearningContinuous Control

Two-Timescale Hierarchical Reinforcement Learning for Resilient Operations

2026-07-26 · Young Hyun Cho, Franz Stoll, Will Wei Sun, Guang Lin 외 arxiv

Unexpected shocks recur in global operations, requiring decision rules that adapt as market and operating conditions change. Many operational systems also have hierarchical structures in which long-term and short-term de…

Hierarchical Reinforcement Learning

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC

2026-07-17 · Ruochen Hou, Shiqi Wang, Beom Jun Kim, Hanzhang Fang 외 arxiv

Humanoid navigation in dynamic environments requires long-horizon planning while respecting short-horizon dynamic and safety constraints. Classical visibility-graph planners combined with model predictive control (MPC) c…

Hierarchical Reinforcement Learning

HiFuzz: Hierarchical Reinforcement Learning for Semantic-Aware and Adaptive CPU Fuzzing

2026-07-07 · Ya Wang, Hanwei Fan, Zhenguo Liu, Xiaofeng Zhou 외 arxiv

Modern processor verification struggles to reach deep architectural states due to the inefficiencies of traditional mutation-based fuzzing. We propose HiFuzz, a novel hierarchical reinforcement learning framework that re…

Hierarchical Reinforcement Learning

전체 481편 보기 →