paper-with-me

홈 › Papers

Enhancing Bandwidth Efficiency for Video Motion Transfer Applications using Deep Learning Based Keypoint Prediction

2024-03-17 · Xue Bai, Tasmiah Haque, Sumit Mohan, Yuliang Cai, Byungheon Jeong, Adam Halasz, Srinjoy Das

We propose a deep learning based novel prediction framework for enhanced bandwidth reduction in motion transfer enabled video applications such as video conferencing, virtual reality gaming and privacy preservation for patient health monitoring. To model complex motion, we use the First Order Motion Model (FOMM) that represents dynamic objects using learned keypoints along with their local affine transformations. Keypoints are extracted by a self-supervised keypoint detector and organized in a time series corresponding to the video frames. Prediction of keypoints, to enable transmission using lower frames per second on the source device, is performed using a Variational Recurrent Neural Network (VRNN). The predicted keypoints are then synthesized to video frames using an optical flow estimator and a generator network. This efficacy of leveraging keypoint based representations in conjunction with VRNN based prediction for both video animation and reconstruction is demonstrated on three diverse datasets. For real-time applications, our results show the effectiveness of our proposed architecture by enabling up to 2x additional bandwidth reduction over existing keypoint based video motion transfer frameworks without significantly compromising video quality.

📄 PDF Abstract BibTeX arXiv:2403.11337

Code (0)

등록된 구현이 없습니다.

Tasks

Optical Flow EstimationPredictionTime Series

Similar Papers 제목 키워드 기반

FE-Adapter: Adapting Image-based Emotion Classifiers to Videos

2024-08-05 · Shreyank N Gowda, Boyan Gao, David A. Clifton

Utilizing large pre-trained models for specific tasks has yielded impressive results. However, fully fine-tuning these increasingly large models is becoming prohibitively resource-intensive. This has led to a focus on mo…

Dynamic Facial Expression RecognitionEmotion RecognitionTransfer LearningVideo Emotion Recognition+1

Robust Emotion Recognition from Low Quality and Low Bit Rate Video: A Deep Learning Approach

2017-09-10 · Bowen Cheng, Zhangyang Wang, Zhaobin Zhang, Zhu Li 외

Emotion recognition from facial expressions is tremendously useful, especially when coupled with smart devices and wireless multimedia applications. However, the inadequate network bandwidth often limits the spatial reso…

DecoderEmotion RecognitionSuper-Resolution

Cross-Camera Human Motion Transfer by Time Series Analysis

2021-09-29 · Yaping Zhao, Guanghan Li, Edmund Y. Lam

With advances in optical sensor technology, heterogeneous camera systems are increasingly used for high-resolution (HR) video acquisition and analysis. However, motion transfer across multiple cameras poses challenges. T…

Pose EstimationTime SeriesTime Series Analysis

EfficientMT: Efficient Temporal Adaptation for Motion Transfer in Text-to-Video Diffusion Models

2025-03-25 · Yufei Cai, Hu Han, Yuxiang Wei, Shiguang Shan 외

The progress on generative models has led to significant advances on text-to-video (T2V) generation, yet the motion controllability of generated videos remains limited. Existing motion transfer methods explored the motio…

Video Generation

Busy-Quiet Video Disentangling for Video Classification

2021-03-29 · Guoxi Huang, Adrian G. Bors

In video data, busy motion details from moving regions are conveyed within a specific frequency bandwidth in the frequency domain. Meanwhile, the rest of the frequencies of video data are encoded with quiet information w…

Action ClassificationAction RecognitionAction Recognition In VideosClassification+2