paper-with-me

Papers

Entity Abstraction in Visual Model-Based Reinforcement Learning

2019-10-28 · Rishi Veerapaneni, John D. Co-Reyes, Michael Chang, Michael Janner, Chelsea Finn, Jiajun Wu, Joshua B. Tenenbaum, Sergey Levine

This paper tests the hypothesis that modeling a scene in terms of entities and their local interactions, as opposed to modeling the scene globally, provides a significant benefit in generalizing to physical tasks in a combinatorial space the learner has not encountered before. We present object-centric perception, prediction, and planning (OP3), which to the best of our knowledge is the first fully probabilistic entity-centric dynamic latent variable framework for model-based reinforcement learning that acquires entity representations from raw visual observations without supervision and uses them to predict and plan. OP3 enforces entity-abstraction -- symmetric processing of each entity representation with the same locally-scoped function -- which enables it to scale to model different numbers and configurations of objects from those in training. Our approach to solving the key technical challenge of grounding these entity representations to actual objects in the environment is to frame this variable binding problem as an inference problem, and we develop an interactive inference algorithm that uses temporal continuity and interactive feedback to bind information about object properties to the entity variables. On block-stacking tasks, OP3 generalizes to novel block configurations and more objects than observed during training, outperforming an oracle model that assumes access to object supervision and achieving two to three times better accuracy than a state-of-the-art video prediction model that does not exhibit entity abstraction.

📄 PDF Abstract BibTeX arXiv:1910.12827

Code (1)

jcoreyes/OP3 공식 구현 pytorch

Tasks

modelModel-based Reinforcement LearningObjectObject Discoveryreinforcement-learningReinforcement LearningReinforcement Learning (RL)Variational InferenceVideo Prediction

Similar Papers 제목 키워드 기반

Object Hallucination-Free Reinforcement Unlearning for Vision-Language Models

2026-05-08 · Kaidi Jia, Yujie Lin, Chengyi Yang, Jiayao Ma 외 arxiv

Vision-language models (VLMs) raise growing concerns about privacy, copyright, and bias, motivating machine unlearning to remove sensitive knowledge. However, existing methods primarily fine-tune the language decoder, le…

Object Recognition

Learning Temporal Abstraction with Information-theoretic Constraints for Hierarchical Reinforcement Learning

2019-09-25 · Wenshan Wang, Yaoyu Hu, Sebastian Scherer

Applying reinforcement learning (RL) to real-world problems will require reasoning about action-reward correlation over long time horizons. Hierarchical reinforcement learning (HRL) methods handle this by dividing the ta…

Hierarchical Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Name Translation based on Fine-grained Named Entity Recognition in a Single Language

2016-05-01 · LREC 2016 5 · Kugatsu Sadamitsu, Itsumi Saito, Taichi Katayama, Hisako Asano 외

We propose named entity abstraction methods with fine-grained named entity labels for improving statistical machine translation (SMT). The methods are based on a bilingual named entity recognizer that uses a monolingual …

Machine Translationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Learning Task Informed Abstractions

2021-06-29 · Xiang Fu, Ge Yang, Pulkit Agrawal, Tommi Jaakkola

Current model-based reinforcement learning methods struggle when operating from complex visual scenes due to their inability to prioritize task-relevant features. To mitigate this problem, we propose learning Task Inform…

Model-based Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Learning Task Informed Abstractions

2021-03-09 · ICLR Workshop SSL-RL 2021 5 · Anonymous

Current model-based reinforcement learning methods struggle when operating from complex visual scenes due to their inability to prioritize task-relevant features. To mitigate this problem, we propose learning Task Inform…

Model-based Reinforcement Learningreinforcement-learningReinforcement Learning (RL)