paper-with-me

Papers

Teaching Unknown Objects by Leveraging Human Gaze and Augmented Reality in Human-Robot Interaction

2023-12-12 · Daniel Weber

Robots are becoming increasingly popular in a wide range of environments due to their exceptional work capacity, precision, efficiency, and scalability. This development has been further encouraged by advances in Artificial Intelligence, particularly Machine Learning. By employing sophisticated neural networks, robots are given the ability to detect and interact with objects in their vicinity. However, a significant drawback arises from the underlying dependency on extensive datasets and the availability of substantial amounts of training data for these object detection models. This issue becomes particularly problematic when the specific deployment location of the robot and the surroundings, are not known in advance. The vast and ever-expanding array of objects makes it virtually impossible to comprehensively cover the entire spectrum of existing objects using preexisting datasets alone. The goal of this dissertation was to teach a robot unknown objects in the context of Human-Robot Interaction (HRI) in order to liberate it from its data dependency, unleashing it from predefined scenarios. In this context, the combination of eye tracking and Augmented Reality created a powerful synergy that empowered the human teacher to communicate with the robot and effortlessly point out objects by means of human gaze. This holistic approach led to the development of a multimodal HRI system that enabled the robot to identify and visually segment the Objects of Interest in 3D space. Through the class information provided by the human, the robot was able to learn the objects and redetect them at a later stage. Due to the knowledge gained from this HRI based teaching, the robot's object detection capabilities exhibited comparable performance to state-of-the-art object detectors trained on extensive datasets, without being restricted to predefined classes, showcasing its versatility and adaptability.

📄 PDF Abstract BibTeX arXiv:2312.07638

Code (0)

등록된 구현이 없습니다.

Tasks

object-detectionObject Detection

Similar Papers 제목 키워드 기반

Gazeformer: Scalable, Effective and Fast Prediction of Goal-Directed Human Attention

2023-03-27 · CVPR 2023 1 · Sounak Mondal, Zhibo Yang, Seoyoung Ahn, Dimitris Samaras 외

Predicting human gaze is important in Human-Computer Interaction (HCI). However, to practically serve HCI applications, gaze prediction models must be scalable, fast, and accurate in their spatial and temporal gaze predi…

DecoderGaze PredictionLanguage ModellingPrediction+2

Gaze-based Object Detection in the Wild

2022-03-29 · Daniel Weber, Wolfgang Fuhl, Andreas Zell, Enkelejda Kasneci

In human-robot collaboration, one challenging task is to teach a robot new yet unknown objects enabling it to interact with them. Thereby, gaze can contain valuable information. We investigate if it is possible to detect…

Objectobject-detectionObject Detection

4D Attention: Comprehensive Framework for Spatio-Temporal Gaze Mapping

2021-07-08 · Shuji Oishi, Kenji Koide, Masashi Yokozuka, Atsuhiko Banno

This study presents a framework for capturing human attention in the spatio-temporal domain using eye-tracking glasses. Attention mapping is a key technology for human perceptual activity analysis or Human-Robot Interact…

Visual Localization

Leveraging recent advances in Pre-Trained Language Models forEye-Tracking Prediction

2021-10-09 · Varun Madhavan, Aditya Girish Pawate, Shraman Pal, Abhranil Chandra

Cognitively inspired Natural Language Pro-cessing uses human-derived behavioral datalike eye-tracking data, which reflect the seman-tic representations of language in the humanbrain to augment the neural nets to solve ar…

Imitation Learning with Human Eye Gaze via Multi-Objective Prediction

2021-02-25 · Ravi Kumar Thakur, MD-Nazmus Samin Sunbeam, Vinicius G. Goecks, Ellen Novoseller 외

Approaches for teaching learning agents via human demonstrations have been widely studied and successfully applied to multiple domains. However, the majority of imitation learning work utilizes only behavioral informatio…

Continuous ControlImitation LearningNavigateRobot Manipulation+2