paper-with-me

홈 › Papers

Real-time Addressee Estimation: Deployment of a Deep-Learning Model on the iCub Robot

2023-11-09 · Carlo Mazzola, Francesco Rea, Alessandra Sciutti

Addressee Estimation is the ability to understand to whom a person is talking, a skill essential for social robots to interact smoothly with humans. In this sense, it is one of the problems that must be tackled to develop effective conversational agents in multi-party and unstructured scenarios. As humans, one of the channels that mainly lead us to such estimation is the non-verbal behavior of speakers: first of all, their gaze and body pose. Inspired by human perceptual skills, in the present work, a deep-learning model for Addressee Estimation relying on these two non-verbal features is designed, trained, and deployed on an iCub robot. The study presents the procedure of such implementation and the performance of the model deployed in real-time human-robot interaction compared to previous tests on the dataset used for the training.

📄 PDF Abstract BibTeX arXiv:2311.05334

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Multi-Modal Explainability Approach for Human-Aware Robots in Multi-Party Conversation

2024-05-20 · Iveta Bečková, Štefan Pócoš, Giulia Belgiovine, Marco Matarese 외

The addressee estimation (understanding to whom somebody is talking) is a fundamental task for human activity recognition in multi-party conversation scenarios. Specifically, in the field of human-robot interaction, it b…

Activity RecognitionBinary ClassificationExplainable artificial intelligenceHuman Activity Recognition

To Whom are You Talking? A Deep Learning Model to Endow Social Robots with Addressee Estimation Skills

2023-08-21 · Carlo Mazzola, Marta Romeo, Francesco Rea, Alessandra Sciutti 외

Communicating shapes our social word. For a robot to be considered social and being consequently integrated in our social environment it is fundamental to understand some of the dynamics that rule human-human communicati…

Deep Learning Based Multi-modal Addressee Recognition in Visual Scenes with Utterances

2018-09-12 · Thao Minh Le, Nobuyuki Shimizu, Takashi Miyazaki, Koichi Shinoda

With the widespread use of intelligent systems, such as smart speakers, addressee recognition has become a concern in human-computer interaction, as more and more people expect such systems to understand complicated soci…

MADNet: Maximizing Addressee Deduction Expectation for Multi-Party Conversation Generation

2023-05-22 · Jia-Chen Gu, Chao-Hong Tan, Caiyuan Chu, Zhen-Hua Ling 외

Modeling multi-party conversations (MPCs) with graph neural networks has been proven effective at capturing complicated and graphical information flows. However, existing methods rely heavily on the necessary addressee l…

Bicubic++: Slim, Slimmer, Slimmest -- Designing an Industry-Grade Super-Resolution Network

2023-05-03 · Bahri Batuhan Bilecen, Mustafa Ayazoglu

We propose a real-time and lightweight single-image super-resolution (SR) network named Bicubic++. Despite using spatial dimensions of the input image across the whole network, Bicubic++ first learns quick reversible dow…

4kImage Super-ResolutionSuper-Resolution