paper-with-me

Papers

SimNet: Enabling Robust Unknown Object Manipulation from Pure Synthetic Data via Stereo

2021-06-30 · Thomas Kollar, Michael Laskey, Kevin Stone, Brijen Thananjeyan, Mark Tjersland

Robot manipulation of unknown objects in unstructured environments is a challenging problem due to the variety of shapes, materials, arrangements and lighting conditions. Even with large-scale real-world data collection, robust perception and manipulation of transparent and reflective objects across various lighting conditions remain challenging. To address these challenges we propose an approach to performing sim-to-real transfer of robotic perception. The underlying model, SimNet, is trained as a single multi-headed neural network using simulated stereo data as input and simulated object segmentation masks, 3D oriented bounding boxes (OBBs), object keypoints, and disparity as output. A key component of SimNet is the incorporation of a learned stereo sub-network that predicts disparity. SimNet is evaluated on 2D car detection, unknown object detection, and deformable object keypoint detection and significantly outperforms a baseline that uses a structured light RGB-D sensor. By inferring grasp positions using the OBB and keypoint predictions, SimNet can be used to perform end-to-end manipulation of unknown objects in both easy and hard scenarios using our fleet of Toyota HSR robots in four home environments. In unknown object grasping experiments, the predictions from the baseline RGB-D network and SimNet enable successful grasps of most of the easy objects. However, the RGB-D baseline only grasps 35% of the hard (e.g., transparent) objects, while SimNet grasps 95%, suggesting that SimNet can enable robust manipulation of unknown objects, including transparent objects, in unknown environments.

📄 PDF Abstract BibTeX arXiv:2106.16118

Code (1)

ToyotaResearchInstitute/simnet 공식 구현 pytorch

Tasks

Keypoint DetectionObjectobject-detectionObject DetectionRobot ManipulationSemantic SegmentationTransparent objects

Similar Papers 제목 키워드 기반

NVIDIA SimNet^{TM}: an AI-accelerated multi-physics simulation framework

2020-12-14 · Oliver Hennigh, Susheela Narasimhan, Mohammad Amin Nabian, Akshay Subramaniam 외

We present SimNet, an AI-driven multi-physics simulation framework, to accelerate simulations across a wide range of disciplines in science and engineering. Compared to traditional numerical solvers, SimNet addresses a w…

GPU

Deep SimNets

2015-06-09 · CVPR 2016 6 · Nadav Cohen, Or Sharir, Amnon Shashua

We present a deep layered architecture that generalizes convolutional neural networks (ConvNets). The architecture, called SimNets, is driven by two operators: (i) a similarity function that generalizes inner-product, an…

Online Inertia Parameter Estimation for Unknown Objects Grasped by a Manipulator Towards Space Applications

2025-12-26 · Akiyoshi Uchida, Antonine Richard, Kentaro Uno, Miguel Olivares-Mendez 외 arxiv

Knowing the inertia parameters of a grasped object is crucial for dynamics-aware manipulation, especially in space robotics with free-floating bases. This work addresses the problem of estimating the inertia parameters o…

simNet: Stepwise Image-Topic Merging Network for Generating Detailed and Comprehensive Image Captions

2018-08-27 · EMNLP 2018 10 · Fenglin Liu, Xuancheng Ren, Yuanxin Liu, Houfeng Wang 외

The encode-decoder framework has shown recent success in image captioning. Visual attention, which is good at detailedness, and semantic attention, which is good at comprehensiveness, have been separately proposed to gro…

DecoderImage Captioning

Predicting Stable Configurations for Semantic Placement of Novel Objects

2021-08-26 · Chris Paxton, Chris Xie, Tucker Hermans, Dieter Fox

Human environments contain numerous objects configured in a variety of arrangements. Our goal is to enable robots to repose previously unseen objects according to learned semantic relationships in novel environments. We …

Motion Planningvalid