paper-with-me

Papers

IRIS: Implicit Reinforcement without Interaction at Scale for Learning Control from Offline Robot Manipulation Data

2019-11-13 · Ajay Mandlekar, Fabio Ramos, Byron Boots, Silvio Savarese, Li Fei-Fei, Animesh Garg, Dieter Fox

Learning from offline task demonstrations is a problem of great interest in robotics. For simple short-horizon manipulation tasks with modest variation in task instances, offline learning from a small set of demonstrations can produce controllers that successfully solve the task. However, leveraging a fixed batch of data can be problematic for larger datasets and longer-horizon tasks with greater variations. The data can exhibit substantial diversity and consist of suboptimal solution approaches. In this paper, we propose Implicit Reinforcement without Interaction at Scale (IRIS), a novel framework for learning from large-scale demonstration datasets. IRIS factorizes the control problem into a goal-conditioned low-level controller that imitates short demonstration sequences and a high-level goal selection mechanism that sets goals for the low-level and selectively combines parts of suboptimal solutions leading to more successful task completions. We evaluate IRIS across three datasets, including the RoboTurk Cans dataset collected by humans via crowdsourcing, and show that performant policies can be learned from purely offline learning. Additional results at https://sites.google.com/stanford.edu/iris/ .

📄 PDF Abstract BibTeX arXiv:1911.05321

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityRobot Manipulation

Similar Papers 제목 키워드 기반

Learning Dynamic User Personas from Implicit Interaction Streams via Iterative Refinement

2026-07-29 · Haifeng Wu arxiv

Personalizing large language models (LLMs) to individual users is essential for improving user experience, yet existing approaches typically rely on explicit preference supervision such as pairwise comparisons or demogra…

Transformers are Sample-Efficient World Models

2022-09-01 · Vincent Micheli, Eloi Alonso, François Fleuret

Deep reinforcement learning agents are notoriously sample inefficient, which considerably limits their application to real-world problems. Recently, many model-based methods have been designed to address this issue, with…

Atari Games 100kDeep Reinforcement Learningreinforcement-learningReinforcement Learning+1

IRIS: Implicit Reward-Guided Internal Sifting for Mitigating Multimodal Hallucination

2026-02-02 · Yuanshuai Li, Yuping Yan, Jirui Han, Fei Ming 외 arxiv

Hallucination remains a fundamental challenge for Multimodal Large Language Models (MLLMs). While Direct Preference Optimization (DPO) is a key alignment framework, existing approaches often rely heavily on costly extern…

ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications

2025-10-11 · Yuxi Mi, Qiuyang Yuan, Zhizhou Zhong, Xuan Zhao 외 arxiv

Recently, iris recognition is regaining prominence in immersive applications such as extended reality as a means of seamless user identification. This application scenario introduces unique challenges compared to traditi…

IRIS: An Immersive Robot Interaction System

2025-02-05 · Xinkai Jiang, Qihao Yuan, Enes Ulas Dincer, Hongyi Zhou 외

This paper introduces IRIS, an immersive Robot Interaction System leveraging Extended Reality (XR), designed for robot data collection and interaction across multiple simulators, benchmarks, and real-world scenarios. Whi…

MuJoCoUnity