paper-with-me

홈 › Papers

Learning Visually Guided Latent Actions for Assistive Teleoperation

2021-05-02 · Siddharth Karamcheti, Albert J. Zhai, Dylan P. Losey, Dorsa Sadigh

It is challenging for humans -- particularly those living with physical disabilities -- to control high-dimensional, dexterous robots. Prior work explores learning embedding functions that map a human's low-dimensional inputs (e.g., via a joystick) to complex, high-dimensional robot actions for assistive teleoperation; however, a central problem is that there are many more high-dimensional actions than available low-dimensional inputs. To extract the correct action and maximally assist their human controller, robots must reason over their context: for example, pressing a joystick down when interacting with a coffee cup indicates a different action than when interacting with knife. In this work, we develop assistive robots that condition their latent embeddings on visual inputs. We explore a spectrum of visual encoders and show that incorporating object detectors pretrained on small amounts of cheap, easy-to-collect structured data enables i) accurately and robustly recognizing the current context and ii) generalizing control embeddings to new objects and tasks. In user studies with a high-dimensional physical robot arm, participants leverage this approach to perform new tasks with unseen objects. Our results indicate that structured visual representations improve few-shot performance and are subjectively preferred by users.

📄 PDF Abstract BibTeX arXiv:2105.00580

Code (1)

Stanford-ILIAD/vla 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Casper: Inferring Diverse Intents for Assistive Teleoperation with Vision Language Models

2025-06-17 · Huihan Liu, Rutav Shah, Shuijing Liu, Jack Pittenger 외

Assistive teleoperation, where control is shared between a human and a robot, enables efficient and intuitive human-robot collaboration in diverse and unstructured environments. A central challenge in real-world assistiv…

Customized Handling of Unintended Interface Operation in Assistive Robots

2020-07-04 · Deepak Gopinath, Mahdieh Nejati Javaremi, Brenna D. Argall

We present an assistance system that reasons about a human's intended actions during robot teleoperation in order to provide appropriate corrections for unintended behavior. We model the human's physical interaction with…

Conformalized Teleoperation: Confidently Mapping Human Inputs to High-Dimensional Robot Actions

2024-06-11 · Michelle Zhao, Reid Simmons, Henny Admoni, Andrea Bajcsy

Assistive robotic arms often have more degrees-of-freedom than a human teleoperator can control with a low-dimensional input, like a joystick. To overcome this challenge, existing approaches use data-driven methods to le…

Conformal PredictionUncertainty Quantification

AIris: An AI-powered Wearable Assistive Device for the Visually Impaired

2024-05-13 · Dionysia Danai Brilli, Evangelos Georgaras, Stefania Tsilivaki, Nikos Melanitis 외

Assistive technologies for the visually impaired have evolved to facilitate interaction with a complex and dynamic world. In this paper, we introduce AIris, an AI-powered wearable device that provides environmental aware…

Face RecognitionObject Recognition

SAPS: Shared Autonomy for Policy Steering by Blending Teleoperation with a Pretrained VLA

2026-06-14 · Crystal Zhou, Jehan Yang, Douglas J. Weber, Zackory Erickson arxiv

Recent advancements in Vision-Language-Action (VLA) models have demonstrated impressive generalist capabilities in robot manipulation, yet these policies can be brittle under out-of-distribution spatial and semantic pert…

Robot Manipulation