paper-with-me

Papers

FetchBot: Object Fetching in Cluttered Shelves via Zero-Shot Sim2Real

2025-02-25 · Weiheng Liu, Yuxuan Wan, Jilong Wang, Yuxuan Kuang, Xuesong Shi, Haoran Li, Dongbin Zhao, Zhizheng Zhang, He Wang

Object fetching from cluttered shelves is an important capability for robots to assist humans in real-world scenarios. Achieving this task demands robotic behaviors that prioritize safety by minimizing disturbances to surrounding objects, an essential but highly challenging requirement due to restricted motion space, limited fields of view, and complex object dynamics. In this paper, we introduce FetchBot, a sim-to-real framework designed to enable zero-shot generalizable and safety-aware object fetching from cluttered shelves in real-world settings. To address data scarcity, we propose an efficient voxel-based method for generating diverse simulated cluttered shelf scenes at scale and train a dynamics-aware reinforcement learning (RL) policy to generate object fetching trajectories within these scenes. This RL policy, which leverages oracle information, is subsequently distilled into a vision-based policy for real-world deployment. Considering that sim-to-real discrepancies stem from texture variations mostly while from geometric dimensions rarely, we propose to adopt depth information estimated by full-fledged depth foundation models as the input for the vision-based policy to mitigate sim-to-real gap. To tackle the challenge of limited views, we design a novel architecture for learning multi-view representations, allowing for comprehensive encoding of cluttered shelf scenes. This enables FetchBot to effectively minimize collisions while fetching objects from varying positions and depths, ensuring robust and safety-aware operation. Both simulation and real-robot experiments demonstrate FetchBot's superior generalization ability, particularly in handling a broad range of real-world scenarios, includ

📄 PDF Abstract BibTeX arXiv:2502.17894

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

ADOPT Please enter a description about the method here

Similar Papers 제목 키워드 기반

Map Space Belief Prediction for Manipulation-Enhanced Mapping

2025-02-28 · Joao Marcos Correia Marques, Nils Dengler, Tobias Zaenker, Jesper Mucke 외

Searching for objects in cluttered environments requires selecting efficient viewpoints and manipulation actions to remove occlusions and reduce uncertainty in object locations, shapes, and categories. In this work, we a…

Decision Making Under UncertaintyPrediction

GraspClutter6D: A Large-scale Real-world Dataset for Robust Perception and Grasping in Cluttered Scenes

2025-04-09 · Seunghyeok Back, Joosoon Lee, KangMin Kim, Heeseon Rho 외

Robust grasping in cluttered environments remains an open challenge in robotics. While benchmark datasets have significantly advanced deep learning methods, they mainly focus on simplistic scenes with light occlusion and…

Pose Estimation

MVRackLay: Monocular Multi-View Layout Estimation for Warehouse Racks and Shelves

2022-11-30 · Pranjali Pathre, Anurag Sahu, Ashwin Rao, Avinash Prabhu 외

In this paper, we propose and showcase, for the first time, monocular multi-view layout estimation for warehouse racks and shelves. Unlike typical layout estimation methods, MVRackLay estimates multi-layered layouts, whe…

TetraGrip: Sensor-Driven Multi-Suction Reactive Object Manipulation in Cluttered Scenes

2025-03-12 · Paolo Torrado, Joshua Levin, Markus Grotz, Joshua Smith

Warehouse robotic systems equipped with vacuum grippers must reliably grasp a diverse range of objects from densely packed shelves. However, these environments present significant challenges, including occlusions, divers…

Object

COSNet: A Novel Semantic Segmentation Network using Enhanced Boundaries in Cluttered Scenes

2024-10-31 · Muhammad Ali, Mamoona Javaid, Mubashir Noman, Mustansar Fiaz 외

Automated waste recycling aims to efficiently separate the recyclable objects from the waste by employing vision-based systems. However, the presence of varying shaped objects having different material types makes it a c…

SegmentationSemantic Segmentation