paper-with-me

Papers

Im2Flow: Motion Hallucination from Static Images for Action Recognition

2017-12-12 · CVPR 2018 6 · Ruohan Gao, Bo Xiong, Kristen Grauman

Existing methods to recognize actions in static images take the images at their face value, learning the appearances---objects, scenes, and body poses---that distinguish each action class. However, such models are deprived of the rich dynamic structure and motions that also define human activity. We propose an approach that hallucinates the unobserved future motion implied by a single snapshot to help static-image action recognition. The key idea is to learn a prior over short-term dynamics from thousands of unlabeled videos, infer the anticipated optical flow on novel static images, and then train discriminative models that exploit both streams of information. Our main contributions are twofold. First, we devise an encoder-decoder convolutional neural network and a novel optical flow encoding that can translate a static image into an accurate flow map. Second, we show the power of hallucinated flow for recognition, successfully transferring the learned motion into a standard two-stream network for activity recognition. On seven datasets, we demonstrate the power of the approach. It not only achieves state-of-the-art accuracy for dense optical flow prediction, but also consistently enhances recognition of actions and dynamic scenes.

📄 PDF Abstract BibTeX arXiv:1712.04109

Code (4)

YashNita/Separate-Object-Sounds-by-Watching-Unlabeled-Video-PyTorch- pytorch
rhgao/Deep-MIML-Network pytorch
rhgao/Im2Flow torch
rhgao/separating-object-sounds pytorch

Tasks

Action RecognitionActivity RecognitionDecoderHallucinationOptical Flow EstimationTemporal Action Localization

Similar Papers 제목 키워드 기반

Making a Case for Learning Motion Representations with Phase

2016-09-06 · S. L. Pintea, J. C. van Gemert

This work advocates Eulerian motion representation learning over the current standard Lagrangian optical flow model. Eulerian motion is well captured by using phase, as obtained by decomposing the image through a complex…

Action Recognitionmotion predictionOptical Flow EstimationRepresentation Learning+1

TransFlow: Motion Knowledge Transfer from Video Diffusion Models to Video Salient Object Detection

2025-07-26 · Suhwan Cho, Minhyeok Lee, Jungho Lee, Sunghun Yang 외 arxiv

Video salient object detection (SOD) relies on motion cues to distinguish salient objects from backgrounds, but training such models is limited by scarce video datasets compared to abundant image datasets. Existing appro…

Video Salient Object Detection

Dynamic Structured Illumination Microscopy with a Neural Space-time Model

2022-06-03 · Ruiming Cao, Fanglin Linda Liu, Li-Hao Yeh, Laura Waller

Structured illumination microscopy (SIM) reconstructs a super-resolved image from multiple raw images captured with different illumination patterns; hence, acquisition speed is limited, making it unsuitable for dynamic s…

Super-Resolution

Flow Dynamics Correction for Action Recognition

2023-10-16 · Lei Wang, Piotr Koniusz

Various research studies indicate that action recognition performance highly depends on the types of motions being extracted and how accurate the human actions are represented. In this paper, we investigate different opt…

Action RecognitionFine-grained Action RecognitionHallucinationOptical Flow Estimation

Transforming Static Images Using Generative Models for Video Salient Object Detection

2024-11-21 · Suhwan Cho, Minhyeok Lee, Jungho Lee, Sangyoun Lee

In many video processing tasks, leveraging large-scale image datasets is a common strategy, as image data is more abundant and facilitates comprehensive knowledge transfer. A typical approach for simulating video from st…

object-detectionObject DetectionSalient Object DetectionTransfer Learning+1