paper-with-me

Continuous Control

73개 벤치마크 · 논문 1,431편 · 이 태스크의 논문 보기 →

Benchmarks

PyBullet Ant

결과 24개

PyBullet HalfCheetah

결과 24개

PyBullet Hopper

결과 24개

PyBullet Walker2D

결과 24개

cartpole.swingup

결과 6개

cheetah.run

결과 6개

finger.turn_hard

결과 6개

walker.stand

결과 6개

walker.walk

결과 6개

2D Walker

결과 3개

Acrobot

결과 3개

Ant

결과 3개

Ant + Gathering

결과 3개

Ant + Maze

결과 3개

Cart-Pole Balancing

결과 3개

Full Humanoid

결과 3개

Half-Cheetah

결과 3개

Hopper

결과 3개

Inverted Pendulum

결과 3개

Mountain Car

결과 3개

Simple Humanoid

결과 3개

Swimmer

결과 3개

Swimmer + Gathering

결과 3개

Swimmer + Maze

결과 3개

acrobot.swingup

결과 3개

ball_in_cup.catch

결과 3개

cartpole.balance

결과 3개

finger.spin

결과 3개

finger.turn_easy

결과 3개

fish.swim

결과 3개

hopper.hop

결과 3개

hopper.stand

결과 3개

humanoid.run

결과 3개

pendulum.swingup

결과 3개

quadruped.run

결과 3개

quadruped.walk

결과 3개

reacher.easy

결과 3개

reacher.hard

결과 3개

walker.run

결과 3개

Most implemented

Proximal Policy Optimization Algorithms

2017-07-20 · 구현 188개

Papers

Controllable Affective Generation via Latent Vector Steering

2026-08-26 · Xixian Yong, Siyuan Chang, Yingying Zhang, Xian Wu 외 arxiv

Large Language Models (LLMs) often produce emotionally flattened responses after alignment, limiting their effectiveness in affect-sensitive applications. In this paper, we propose EmoVec, a lightweight framework for con…

Continuous Control

Graph-Operator World Models for Morphology-Parameter Generalization in Continuous Control

2026-08-21 · Xu Yang, Yiqin Yang, Qianchuan Zhao arxiv

World models for continuous control are commonly trained for a fixed physical system and can degrade when known morphology parameters such as link lengths, masses, damping, and actuation change. Existing approaches often…

Continuous Control

Orthogonal JEPA: Factorized Predictive States for Latent World Models

2026-08-20 · Taoyong Cui, Pheng Ann Heng, Wanli Ouyang arxiv

World models construct latent states that support prediction, planning, and reasoning about an underlying system. Joint-embedding predictive architectures (JEPAs) offer a direct way to learn such states by predicting tar…

Continuous Control

To Go Far, Go Together: Diverse Preferences Induce a Curriculum for Reward Optimization

2026-08-19 · Taehyung Kim, Jongeun Choi arxiv

Learning a reward model from human feedback and optimizing a policy against it is one approach to aligning AI systems with individual users. From a fairness perspective, existing work improves such alignment by developin…

Continuous Control

An Omitted Mode Is a Rare Rule: The Sampling-Verification Danger Law in Continuous Code World Models

2026-08-18 · Javier Aguilar Martín arxiv

In the Code World Model paradigm an LLM synthesizes an executable world model that a classical planner searches, and the model is accepted when it reproduces sampled transitions. We ask what that acceptance certifies in …

Continuous Control

Observation-Grounded Self-Predictive Reinforcement Learning for Visual Continuous Control

2026-08-06 · Xinwei Liu, Junyuan Liang, Jianting Zhang, Wuhui Chen arxiv

Sample-efficient policy learning from pixels is a long-standing challenge in reinforcement learning (RL). Recent dynamics-based representation learning methods have significantly improved the sample efficiency of model-f…

Representation LearningReinforcement LearningContinuous Control

전체 1,431편 보기 →