paper-with-me

Papers

Self-Supervised Learning of Action Affordances as Interaction Modes

2023-05-27 · Liquan Wang, Nikita Dvornik, Rafael Dubeau, Mayank Mittal, Animesh Garg

When humans perform a task with an articulated object, they interact with the object only in a handful of ways, while the space of all possible interactions is nearly endless. This is because humans have prior knowledge about what interactions are likely to be successful, i.e., to open a new door we first try the handle. While learning such priors without supervision is easy for humans, it is notoriously hard for machines. In this work, we tackle unsupervised learning of priors of useful interactions with articulated objects, which we call interaction modes. In contrast to the prior art, we use no supervision or privileged information; we only assume access to the depth sensor in the simulator to learn the interaction modes. More precisely, we define a successful interaction as the one changing the visual environment substantially and learn a generative model of such interactions, that can be conditioned on the desired goal state of the object. In our experiments, we show that our model covers most of the human interaction modes, outperforms existing state-of-the-art methods for affordance learning, and can generalize to objects never seen during training. Additionally, we show promising results in the goal-conditional setup, where our model can be quickly fine-tuned to perform a given task. We show in the experiments that such affordance learning predicts interaction which covers most modes of interaction for the querying articulated object and can be fine-tuned to a goal-conditional model. For supplementary: https://actaim.github.io.

📄 PDF Abstract BibTeX arXiv:2305.17565

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectSelf-Supervised Learning

Similar Papers 제목 키워드 기반

Pretraining over Interactions for Learning Grounded Object Representations

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Large language models have been criticized for their limited ability to reason about \textit{affordances} - the actions that can be performed on an object. It has been argued that to accomplish this, models need some for…

Object

Object-agnostic Affordance Categorization via Unsupervised Learning of Graph Embeddings

2023-03-30 · Alexia Toumpa, Anthony G. Cohn

Acquiring knowledge about object interactions and affordances can facilitate scene understanding and human-robot collaboration tasks. As humans tend to use objects in many different ways depending on the scene and the ob…

ObjectScene Understanding

Grounded Human-Object Interaction Hotspots from Video

2018-12-11 · ICCV 2019 10 · Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman

Learning how to interact with objects is an important step towards embodied visual intelligence, but existing techniques suffer from heavy supervision or sensing requirements. We propose an approach to learn human-object…

Human-Object Interaction DetectionObjectObject RecognitionSemantic Segmentation+1

Grounded Human-Object Interaction Hotspots from Video (Extended Abstract)

2019-06-03 · Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman

Learning how to interact with objects is an important step towards embodied visual intelligence, but existing techniques suffer from heavy supervision or sensing requirements. We propose an approach to learn human-object…

Human-Object Interaction DetectionObjectSemantic Segmentation

Wordcraft: a Human-AI Collaborative Editor for Story Writing

2021-07-15 · Andy Coenen, Luke Davis, Daphne Ippolito, Emily Reif 외

As neural language models grow in effectiveness, they are increasingly being applied in real-world settings. However these applications tend to be limited in the modes of interaction they support. In this extended abstra…

Few-Shot Learning