paper-with-me

홈 › Papers

Learning End-to-End Action Interaction by Paired-Embedding Data Augmentation

2020-07-16 · Ziyang Song, Zejian yuan, Chong Zhang, Wanchao Chi, Yonggen Ling, Shenghao Zhang

In recognition-based action interaction, robots' responses to human actions are often pre-designed according to recognized categories and thus stiff. In this paper, we specify a new Interactive Action Translation (IAT) task which aims to learn end-to-end action interaction from unlabeled interactive pairs, removing explicit action recognition. To enable learning on small-scale data, we propose a Paired-Embedding (PE) method for effective and reliable data augmentation. Specifically, our method first utilizes paired relationships to cluster individual actions in an embedding space. Then two actions originally paired can be replaced with other actions in their respective neighborhood, assembling into new pairs. An Act2Act network based on conditional GAN follows to learn from augmented data. Besides, IAT-test and IAT-train scores are specifically proposed for evaluating methods on our task. Experimental results on two datasets show impressive effects and broad application prospects of our method.

📄 PDF Abstract BibTeX arXiv:2007.08071

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionData AugmentationTranslation

Similar Papers 제목 키워드 기반

INTRA: Interaction Relationship-aware Weakly Supervised Affordance Grounding

2024-09-10 · Ji Ha Jang, Hoigi Seo, Se Young Chun

Affordance denotes the potential interactions inherent in objects. The perception of affordance can enable intelligent agents to navigate and interact with new environments efficiently. Weakly supervised affordance groun…

Contrastive LearningLanguage ModelingLanguage ModellingNavigate+1

Learning Reactive Human Motion Generation from Paired Interaction Data Using Transformer-Based Models

2026-04-24 · Masato Soga, Ryuki Takebayashi arxiv

Recent advances in deep learning have enabled the generation of videos from textual descriptions as well as the prediction of future sequences from input videos. Similarly, in human motion modeling, motions can be genera…

Wearable Sensor-Based Few-Shot Continual Learning on Hand Gestures for Motor-Impaired Individuals via Latent Embedding Exploitation

2024-05-14 · Riyad Bin Rafiq, Weishi Shi, Mark V. Albert

Hand gestures can provide a natural means of human-computer interaction and enable people who cannot speak to communicate efficiently. Existing hand gesture recognition methods heavily depend on pre-defined gestures, how…

Continual LearningGesture RecognitionHand Gesture RecognitionHand-Gesture Recognition

Controlling Multimodal Conversational Agents with Coverage-Enhanced Latent Actions

2026-01-12 · Yongqi Li, Hao Lang, Tieyun Qian, Yongbin Li arxiv

Vision-language models are increasingly employed as multimodal conversational agents (MCAs) for diverse conversational tasks. Recently, reinforcement learning (RL) has been widely explored for adapting MCAs to various hu…

Reinforcement Learning

Bringing Cognitive Augmentation to Web Browsing Accessibility

2020-12-07 · Alessandro Pina, Marcos Baez, Florian Daniel

In this paper we explore the opportunities brought by cognitive augmentation to provide a more natural and accessible web browsing experience. We explore these opportunities through \textit{conversational web browsing}, …