paper-with-me

홈 › Papers

A Learning Approach to Robot-Agnostic Force-Guided High Precision Assembly

2020-10-15 · Jieliang Luo, Hui Li

In this work we propose a learning approach to high-precision robotic assembly problems. We focus on the contact-rich phase, where the assembly pieces are in close contact with each other. Unlike many learning-based approaches that heavily rely on vision or spatial tracking, our approach takes force/torque in task space as the only observation. Our training environment is robotless, as the end-effector is not attached to any specific robot. Trained policies can then be applied to different robotic arms without re-training. This approach can greatly reduce complexity to perform contact-rich robotic assembly in the real world, especially in unstructured settings such as in architectural construction. To achieve it, we have developed a new distributed RL agent, named Recurrent Distributed DDPG (RD2), which extends Ape-X DDPG with recurrency and makes two structural improvements on prioritized experience replay. Our results show that RD2 is able to solve two fundamental high-precision assembly tasks, lap-joint and peg-in-hole, and outperforms two state-of-the-art algorithms, Ape-X DDPG and PPO with LSTM. We have successfully evaluated our robot-agnostic policies on three robotic arms, Kuka KR60, Franka Panda, and UR10, in simulation. The video presenting our experiments is available at https://sites.google.com/view/rd2-rl

📄 PDF Abstract BibTeX arXiv:2010.08052

Code (0)

등록된 구현이 없습니다.

Tasks

OpenAI GymVocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

Batch Normalization 설명 없음
Weight Decay 설명 없음
Entropy Regularization 설명 없음
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Adam 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
PPO Proximal Policy Optimization, or PPO, is a policy gradient method for reinforcement learning. The motivation was to have an algorithm with the data efficiency and reliable…

Similar Papers 제목 키워드 기반

Keyframe-Guided Structured Rewards for Reinforcement Learning in Long-Horizon Laboratory Robotics

2026-02-28 · Yibo Qiu, Shu'ang Sun, Haoliang Ye, Ronald X Xu 외 arxiv

Long-horizon precision manipulation in laboratory automation, such as pipette tip attachment and liquid transfer, requires policies that respect strict procedural logic while operating in continuous, high-dimensional sta…

Reinforcement Learning

Force-Based Viscosity and Elasticity Measurements for Material Biomechanical Characterisation with a Collaborative Robotic Arm

2025-07-15 · Luca Beber, Edoardo Lamon, Giacomo Moretti, Matteo Saveriano 외 arxiv

Diagnostic activities, such as ultrasound scans and palpation, are relatively low-cost. They play a crucial role in the early detection of health problems and in assessing their progression. However, they are also error-…

Resolution-Enhanced MRI-Guided Navigation of Spinal Cellular Injection Robot

2020-06-09 · Daniel Enrique Martinez, Waiman Meinhold, John Oshinski, Ai-Ping Hu 외

This paper presents a method of navigating a surgical robot beyond the resolution of magnetic resonance imaging (MRI) by using a resolution enhancement technique enabled by high-precision piezoelectric actuation. The sur…

Navigate

Reliability-aware Execution Gating for Near-field and Off-axis Vision-guided Robotic Alignment

2026-02-09 · Ning Hu, Senhao Cao, Maochen Li arxiv

Vision-guided robotic systems are increasingly deployed in precision alignment tasks that require reliable execution under near-field and off-axis configurations. While recent advances in pose estimation have significant…

Pose Estimation

HALO-WA: Hybrid-Attention Latent-Guided Online Reinforcement Learning for World-Action Models

2026-07-05 · Angen Ye, Weijie Ke, Xiaofeng Wang, Xinze Chen 외 arxiv

World-action (WA) models can generate long-horizon action chunks for general-purpose robotic manipulation, but they remain vulnerable to calibration, perception, and contact-dynamics errors in real-world precision tasks,…

Reinforcement Learning