paper-with-me

홈 › Papers

OpenRoboCare: A Multimodal Multi-Task Expert Demonstration Dataset for Robot Caregiving

2025-11-17 · Xiaoyu Liang, Ziang Liu, Kelvin Lin, Edward Gu, Ruolin Ye, Tam Nguyen, Cynthia Hsu, Zhanxin Wu, Xiaoman Yang, Christy Sum Yu Cheung, Harold Soh, Katherine Dimitropoulou, Tapomayukh Bhattacharjee arxiv

We present OpenRoboCare, a multimodal dataset for robot caregiving, capturing expert occupational therapist demonstrations of Activities of Daily Living (ADLs). Caregiving tasks involve complex physical human-robot interactions, requiring precise perception under occlusions, safe physical contact, and long-horizon planning. While recent advances in robot learning from demonstrations have shown promise, there is a lack of a large-scale, diverse, and expert-driven dataset that captures real-world caregiving routines. To address this gap, we collect data from 21 occupational therapists performing 15 ADL tasks on two manikins. The dataset spans five modalities: RGB-D video, pose tracking, eye-gaze tracking, task and action annotations, and tactile sensing, providing rich multimodal insights into caregiver movement, attention, force application, and task execution strategies. We further analyze expert caregiving principles and strategies, offering insights to improve robot efficiency and task feasibility. Additionally, our evaluations demonstrate that OpenRoboCare presents challenges for state-of-the-art robot perception and human activity recognition methods, both critical for developing safe and adaptive assistive robots, highlighting the value of our contribution. See our website for additional visualizations: https://emprise.cs.cornell.edu/robo-care/.

📄 PDF Abstract BibTeX arXiv:2511.13707

Code (0)

등록된 구현이 없습니다.

Tasks

Human Activity RecognitionPose Tracking

Similar Papers 제목 키워드 기반

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

2024-12-02 · Anian Ruoss, Fabio Pardo, Harris Chan, Bonnie Li 외

Today's largest foundation models have increasingly general capabilities, yet when used as agents, they often struggle with simple reasoning and decision-making tasks, even though they possess good factual knowledge of t…

Decision MakingImitation Learning

Imitation Learning for Fashion Style Based on Hierarchical Multimodal Representation

2020-04-13 · Shizhu Liu, Shanglin Yang, Hui Zhou

Fashion is a complex social phenomenon. People follow fashion styles from demonstrations by experts or fashion icons. However, for machine agent, learning to imitate fashion experts from demonstrations can be challenging…

Imitation LearningReinforcement Learning

ExpertAF: Expert Actionable Feedback from Video

2024-08-01 · CVPR 2025 1 · Kumar Ashutosh, Tushar Nagarajan, Georgios Pavlakos, Kris Kitani 외

Feedback is essential for learning a new skill or improving one's current skill-level. However, current methods for skill-assessment from video only provide scores or compare demonstrations, leaving the burden of knowing…

Language ModelingLanguage ModellingVideo Retrieval

Hierarchical Imitation Learning of Team Behavior from Heterogeneous Demonstrations

2025-02-24 · Sangwon Seo, Vaibhav Unhelkar

Successful collaboration requires team members to stay aligned, especially in complex sequential tasks. Team members must dynamically coordinate which subtasks to perform and in what order. However, real-world constraint…

Imitation Learning

Guide Your Agent with Adaptive Multimodal Rewards

2023-09-21 · NeurIPS 2023 11

Developing an agent capable of adapting to unseen environments remains a difficult challenge in imitation learning. This work presents Adaptive Return-conditioned Policy (ARP), an efficient framework designed to enhance …