paper-with-me

홈 › Papers

Enhancing Robot Learning through Learned Human-Attention Feature Maps

2023-08-29 · Daniel Scheuchenstuhl, Stefan Ulmer, Felix Resch, Luigi Berducci, Radu Grosu

Robust and efficient learning remains a challenging problem in robotics, in particular with complex visual inputs. Inspired by human attention mechanism, with which we quickly process complex visual scenes and react to changes in the environment, we think that embedding auxiliary information about focus point into robot learning would enhance efficiency and robustness of the learning process. In this paper, we propose a novel approach to model and emulate the human attention with an approximate prediction model. We then leverage this output and feed it as a structured auxiliary feature map into downstream learning tasks. We validate this idea by learning a prediction model from human-gaze recordings of manual driving in the real world. We test our approach on two learning tasks - object detection and imitation learning. Our experiments demonstrate that the inclusion of predicted human attention leads to improved robustness of the trained models to out-of-distribution samples and faster learning in low-data regime settings. Our work highlights the potential of incorporating structured auxiliary information in representation learning for robotics and opens up new avenues for research in this direction. All code and data are available online.

📄 PDF Abstract BibTeX arXiv:2308.15327

Code (1)

cps-tuwien/learning_human_attention 공식 구현

Tasks

Imitation Learningobject-detectionObject DetectionRepresentation Learning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Gaze-Regularized Vision-Language-Action Models for Robotic Manipulation

2026-03-24 · Anupam Pani, Yanchao Yang arxiv

Despite advances in Vision-Language-Action (VLA) models, robotic manipulation struggles with fine-grained tasks because current models lack mechanisms for active visual attention allocation. Human gaze naturally encodes …

Show, Attend and Interact: Perceivable Human-Robot Social Interaction through Neural Attention Q-Network

2017-02-28 · Ahmed Hussain Qureshi, Yutaka Nakamura, Yuichiro Yoshikawa, Hiroshi Ishiguro

For a safe, natural and effective human-robot social interaction, it is essential to develop a system that allows a robot to demonstrate the perceivable responsive behaviors to complex human behaviors. We introduce the M…

Deep Attentionreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition

2024-06-29 · Shengcheng Luo, Quanquan Peng, Jun Lv, Kaiwen Hong 외

Employing a teleoperation system for gathering demonstrations offers the potential for more efficient learning of robot manipulation. However, teleoperating a robot arm equipped with a dexterous hand or gripper, via a te…

Robot Manipulation

Speech Driven Backchannel Generation using Deep Q-Network for Enhancing Engagement in Human-Robot Interaction

2019-08-05 · Nusrah Hussain, Engin Erzin, T. Metin Sezgin, Yucel Yemez

We present a novel method for training a social robot to generate backchannels during human-robot interaction. We address the problem within an off-policy reinforcement learning framework, and show how a robot may learn …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Awakening Facial Emotional Expressions in Human-Robot

2025-10-27 · Yongtong Zhu, Lei Li, Iggy Qian, WenBin Zhou 외 arxiv

The facial expression generation capability of humanoid social robots is critical for achieving natural and human-like interactions, playing a vital role in enhancing the fluidity of human-robot interactions and the accu…