paper-with-me

Papers

Affordances from Human Videos as a Versatile Representation for Robotics

2023-04-17 · CVPR 2023 1 · Shikhar Bahl, Russell Mendonca, Lili Chen, Unnat Jain, Deepak Pathak

Building a robot that can understand and learn to interact by watching humans has inspired several vision problems. However, despite some successful results on static datasets, it remains unclear how current models can be used on a robot directly. In this paper, we aim to bridge this gap by leveraging videos of human interactions in an environment centric manner. Utilizing internet videos of human behavior, we train a visual affordance model that estimates where and how in the scene a human is likely to interact. The structure of these behavioral affordances directly enables the robot to perform many complex tasks. We show how to seamlessly integrate our affordance model with four robot learning paradigms including offline imitation learning, exploration, goal-conditioned learning, and action parameterization for reinforcement learning. We show the efficacy of our approach, which we call VRB, across 4 real world environments, over 10 different tasks, and 2 robotic platforms operating in the wild. Results, visualizations and videos at https://robo-affordances.github.io/

📄 PDF Abstract BibTeX arXiv:2304.08488

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation Learning

Similar Papers 제목 키워드 기반

Look Ma, No Hands! Agent-Environment Factorization of Egocentric Videos

2023-05-25 · NeurIPS 2023 11

The analysis and use of egocentric videos for robotic tasks is made challenging by occlusion due to the hand and the visual mismatch between the human hand and a robot end-effector. In this sense, the human hand presents…

3D ReconstructionObjectobject-detectionObject Detection+1

Contextual Affordances for Safe Exploration in Robotic Scenarios

2024-05-10 · William Z. Ye, Eduardo B. Sandoval, Pamela Carreno-Medrano, Francisco Cru

Robotics has been a popular field of research in the past few decades, with much success in industrial applications such as manufacturing and logistics. This success is led by clearly defined use cases and controlled ope…

Safe Exploration

RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

2024-11-05 · Soroush Nasiriany, Sean Kirmani, Tianli Ding, Laura Smith 외

We explore how intermediate policy representations can facilitate generalization by providing guidance on how to perform manipulation tasks. Existing representations such as language, goal images, and trajectory sketches…

Robot Manipulation

Agent-Exploitation Affordances: From Basic to Complex Representation Patterns

2026-07-08 · Bastien Dussard, Aurélie Clodic, Guillaume Sarthou arxiv

In robotics, the capability of an artificial agent to represent the range of its action possibilities, i.e. affordances, is crucial to understand how it can act on its environment. While functional affordances, which ref…

Demo2Vec: Reasoning Object Affordances From Online Videos

2018-06-01 · CVPR 2018 6 · Kuan Fang, Te-Lin Wu, Daniel Yang, Silvio Savarese 외

Watching expert demonstrations is an important way for humans and robots to reason about affordances of unseen objects. In this paper, we consider the problem of reasoning object affordances through the feature embedding…

ObjectVideo-to-image Affordance Grounding