paper-with-me

홈 › Papers

Rainbow Deep Q-Learning with Kinematics-Aware Design for Cooperative Delta and 3-RRS Parallel Robot Insertion

2026-05-12 · Hassen Nigatu, Gaokun Shi, Jituo Li, Wang Jin, Lu Guodong arxiv

This paper presents a kinematics-aware deep reinforcement learning framework based on Rainbow Deep Q-Networks (DQN) for cooperative peg-in-hole manipulation by a Delta parallel robot and a 3-RRS (Revolute--Revolute--Spherical) parallel manipulator. A key contribution is the integration of a geometric design-optimization stage that precedes learning: the 3-RRS geometry is tuned to maximize the singularity-free workspace and improve conditioning, which in turn enlarges the safe region in which the reinforcement learning policy can explore. Together the two manipulators expose a 6~degree-of-freedom (DoF) controllable subspace (three Delta translations, two 3-RRS rotations, and one 3-RRS vertical translation); the peg-in-hole task is invariant to rotation about the peg axis, so the task-relevant manifold is five dimensional. The cooperative insertion problem is cast as a Markov Decision Process with a 12-dimensional state vector and a discrete action set containing $6 \times 2 = 12$ incremental commands (one positive and one negative per controlled DoF). A shaped reward combines dense proximity guidance, penalties for kinematic and workspace violations, and sparse bonuses for successful insertions. The Rainbow DQN -- integrating double Q-learning, dueling architecture, prioritized replay, multi-step returns, noisy linear layers for exploration, and a distributional value head -- is trained with a two-stage curriculum. The co-designed framework is validated in a high-fidelity kinematic simulator, where it achieves stable policy convergence, reliable insertions, and reduced constraint violations compared against a vanilla DQN agent and a classical sampling-based planner.

📄 PDF Abstract BibTeX arXiv:2605.11697

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Generalized Rainbow Differential Privacy

2023-09-11 · Yuzhou Gu, Ziqi Zhou, Onur Günlü, Rafael G. L. D'Oliveira 외

We study a new framework for designing differentially private (DP) mechanisms via randomized graph colorings, called rainbow differential privacy. In this framework, datasets are nodes in a graph, and two neighboring dat…

valid

Kinematics-Aware Multi-Policy Reinforcement Learning for Force-Capable Humanoid Loco-Manipulation

2025-11-26 · Kaiyan Xiao, Zihan Xu, Cheng Zhe, Chengju Liu 외 arxiv

Humanoid robots, with their human-like morphology, hold great potential for industrial applications. However, existing loco-manipulation methods primarily focus on dexterous manipulation, falling short of the combined re…

Reinforcement Learning

Evaluating the Rainbow DQN Agent in Hanabi with Unseen Partners

2020-04-28 · Rodrigo Canaan, Xianbo Gao, Youjin Chung, Julian Togelius 외

Hanabi is a cooperative game that challenges exist-ing AI techniques due to its focus on modeling the mental states ofother players to interpret and predict their behavior. While thereare agents that can achieve near-per…

Spectral functions in Minkowski quantum electrodynamics from neural reconstruction

2025-10-06 · Rodrigo Carmo Terin arxiv

We study neural reconstructions of quenched rainbow quantum electrodynamics (QED) Dyson--Schwinger benchmarks in Minkowski-related kinematics. Using the dispersive formulation as motivation, we separate the Euclidean Fuk…

Distributed Bayesian Online Learning for Cooperative Manipulation

2021-04-09 · Pablo Budde gen. Dohmann, Armin Lederer, Marcel Dißemond, Sandra Hirche

For tasks where the dynamics of multiple agents are physically coupled, e.g., in cooperative manipulation, the coordination between the individual agents becomes crucial, which requires exact knowledge of the interaction…