paper-with-me

홈 › Papers

HERO: Human Reaction Generation from Videos

2025-03-11 · Chengjun Yu, Wei Zhai, Yuhang Yang, Yang Cao, Zheng-Jun Zha

Human reaction generation represents a significant research domain for interactive AI, as humans constantly interact with their surroundings. Previous works focus mainly on synthesizing the reactive motion given a human motion sequence. This paradigm limits interaction categories to human-human interactions and ignores emotions that may influence reaction generation. In this work, we propose to generate 3D human reactions from RGB videos, which involves a wider range of interaction categories and naturally provides information about expressions that may reflect the subject's emotions. To cope with this task, we present HERO, a simple yet powerful framework for Human rEaction geneRation from videOs. HERO considers both global and frame-level local representations of the video to extract the interaction intention, and then uses the extracted interaction intention to guide the synthesis of the reaction. Besides, local visual representations are continuously injected into the model to maximize the exploitation of the dynamic properties inherent in videos. Furthermore, the ViMo dataset containing paired Video-Motion data is collected to support the task. In addition to human-human interactions, these video-motion pairs also cover animal-human interactions and scene-human interactions. Extensive experiments demonstrate the superiority of our methodology. The code and dataset will be publicly available at https://jackyu6.github.io/HERO.

📄 PDF Abstract BibTeX arXiv:2503.08270

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

MuSteerNet: Human Reaction Generation from Videos via Observation-Reaction Mutual Steering

2026-03-20 · Yuan Zhou, Yongzhi Li, Yanqi Dai, Xingyu Zhu 외 arxiv

Video-driven human reaction generation aims to synthesize 3D human motions that directly react to observed video sequences, which is crucial for building human-like interactive AI systems. However, existing methods often…

EgoReAct: Egocentric Video-Driven 3D Human Reaction Generation

2025-12-28 · Libo Zhang, Zekun Li, Tianyu Li, Zeyu Cao 외 arxiv

Humans exhibit adaptive, context-sensitive responses to egocentric visual input. However, faithfully modeling such reactions from egocentric video remains challenging due to the dual requirements of strictly causal gener…

Understanding Video Content: Efficient Hero Detection and Recognition for the Game "Honor of Kings"

2019-07-18 · Wentao Yao, Zixun Sun, Xiao Chen

In order to understand content and automatically extract labels for videos of the game "Honor of Kings", it is necessary to detect and recognize characters (called "hero") together with their camps in the game video. In …

Template Matching

Single Input Multi Output Model of Molecular Communication via Diffusion with Spheroidal Receivers

2024-05-22 · Ibrahim Isik, Mitra Rezaei, Adam Noel

Spheroids are aggregates of cells that can mimic the cellular organization often found in tissues. They are typically formed through the self-assembly of cells in a culture where there is a promotion of interactions and …

HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning

2024-10-07 · Ayano Hiranaka, Shang-Fu Chen, Chieh-Hsin Lai, Dongjun Kim 외

Controllable generation through Stable Diffusion (SD) fine-tuning aims to improve fidelity, safety, and alignment with human guidance. Existing reinforcement learning from human feedback methods usually rely on predefine…

Image Generationreinforcement-learningReinforcement LearningRepresentation Learning