paper-with-me

Papers

DynamoNet: Dynamic Action and Motion Network

2019-04-25 · ICCV 2019 10 · Ali Diba, Vivek Sharma, Luc van Gool, Rainer Stiefelhagen

In this paper, we are interested in self-supervised learning the motion cues in videos using dynamic motion filters for a better motion representation to finally boost human action recognition in particular. Thus far, the vision community has focused on spatio-temporal approaches using standard filters, rather we here propose dynamic filters that adaptively learn the video-specific internal motion representation by predicting the short-term future frames. We name this new motion representation, as dynamic motion representation (DMR) and is embedded inside of 3D convolutional network as a new layer, which captures the visual appearance and motion dynamics throughout entire video clip via end-to-end network learning. Simultaneously, we utilize these motion representation to enrich video classification. We have designed the frame prediction task as an auxiliary task to empower the classification problem. With these overall objectives, to this end, we introduce a novel unified spatio-temporal 3D-CNN architecture (DynamoNet) that jointly optimizes the video classification and learning motion representation by predicting future frames as a multi-task learning problem. We conduct experiments on challenging human action datasets: Kinetics 400, UCF101, HMDB51. The experiments using the proposed DynamoNet show promising results on all the datasets.

📄 PDF Abstract BibTeX arXiv:1904.11407

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionClassificationGeneral ClassificationMulti-Task LearningSelf-Supervised LearningTemporal Action LocalizationVideo Classification

Similar Papers 제목 키워드 기반

Dynamic Full-body Motion Agent with Object Interaction via Blending Pre-trained Modular Controllers

2026-05-12 · Sanghyeok Nam, Byoungjun Kim, Daehyung Park, Tae-Kyun Kim arxiv

Generating physically plausible dynamic motions of human-object interaction (HOI) remains challenging, mainly due to existing HOI datasets limited to static interactions, and pretrained agents capable of either dynamic f…

Fine-grained text-driven dual-human motion generation via dynamic hierarchical interaction

2025-10-09 · Mu Li, Yin Wang, Zhiying Leng, Jiapeng Liu 외 arxiv

Human interaction is inherently dynamic and hierarchical, where the dynamic refers to the motion changes with distance, and the hierarchy is from individual to inter-individual and ultimately to overall motion. Exploitin…

Action-guided 3D Human Motion Prediction

2021-12-01 · NeurIPS 2021 12 · Jiangxin Sun, Zihang Lin, Xintong Han, Jian-Fang Hu 외

The ability of forecasting future human motion is important for human-machine interaction systems to understand human behaviors and make interaction. In this work, we focus on developing models to predict future human mo…

Human motion predictionmotion predictionPrediction

DynamicWAM: Dual-Path Motion Conditioning for World-Action Models in Dynamic Manipulation

2026-08-01 · Yunfan Lou, Hewen Gao, Xiyu Zhu, Zhuoran Qiao 외 arxiv

Dynamic manipulation requires robots to infer target motion and respond promptly, yet existing World-Action Models (WAMs) typically condition only on the current frame and execute large backbones synchronously, limiting …

SoMoFormer: Social-Aware Motion Transformer for Multi-Person Motion Prediction

2022-08-19 · Xiaogang Peng, Yaodi Shen, Haoran Wang, Binling Nie 외

Multi-person motion prediction remains a challenging problem, especially in the joint representation learning of individual motion and social interactions. Most prior methods only involve learning local pose dynamics for…

motion predictionRepresentation Learning