paper-with-me

Papers

Combining Spatio-Temporal Appearance Descriptors and Optical Flow for Human Action Recognition in Video Data

2013-10-01 · Karla Brkić, Srđan Rašić, Axel Pinz, Siniša Šegvić, Zoran Kalafatić

This paper proposes combining spatio-temporal appearance (STA) descriptors with optical flow for human action recognition. The STA descriptors are local histogram-based descriptors of space-time, suitable for building a partial representation of arbitrary spatio-temporal phenomena. Because of the possibility of iterative refinement, they are interesting in the context of online human action recognition. We investigate the use of dense optical flow as the image function of the STA descriptor for human action recognition, using two different algorithms for computing the flow: the Farneb\"ack algorithm and the TVL1 algorithm. We provide a detailed analysis of the influencing optical flow algorithm parameters on the produced optical flow fields. An extensive experimental validation of optical flow-based STA descriptors in human action recognition is performed on the KTH human action dataset. The encouraging experimental results suggest the potential of our approach in online human action recognition.

📄 PDF Abstract BibTeX arXiv:1310.0308

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionOptical Flow EstimationTemporal Action Localization

Similar Papers 제목 키워드 기반

Adaptive Semantic-Spatio-Temporal Graph Convolutional Network for Lip Reading

2021-08-16 · IEEE Transactions on Multimedia 2021 8 · Changchong Sheng, Xinzhong Zhu, Huiying Xu, Matti Pietikäinen 외

The goal of this work is to recognize words, phrases, and sentences being spoken by a talking face without given the audio. Current deep learning approaches for lip reading focus on exploring the appearance and optical f…

Landmark-based LipreadingLip ReadingOptical Flow Estimation

Pooled Motion Features for First-Person Videos

2014-12-19 · CVPR 2015 6 · M. S. Ryoo, Brandon Rothrock, Larry Matthies

In this paper, we present a new feature representation for first-person videos. In first-person video understanding (e.g., activity recognition), it is very important to capture both entire scene dynamics (i.e., egomotio…

Activity RecognitionActivity Recognition In VideosTime SeriesTime Series Analysis+1

Nested Motion Descriptors

2015-06-01 · CVPR 2015 6 · Jeffrey Byrne

A nested motion descriptor is a spatiotemporal representation of motion that is invariant to global camera translation, without requiring an explicit estimate of optical flow or camera stabilization. This descriptor is …

Activity RecognitionOptical Flow EstimationTranslation

Modeling Spatio-Temporal Human Track Structure for Action Localization

2018-06-28 · Guilhem Chéron, Anton Osokin, Ivan Laptev, Cordelia Schmid

This paper addresses spatio-temporal localization of human actions in video. In order to localize actions in time, we propose a recurrent localization network (RecLNet) designed to model the temporal structure of actions…

Action LocalizationHuman DetectionOptical Flow EstimationSpatio-Temporal Action Localization+2

Zigzag persistence for coral reef resilience using a stochastic spatial model

2022-09-19 · Robert A. McDonald, Rosanna Neuhausler, Martin Robinson, Laurel G. Larsen 외

A complex interplay between species governs the evolution of spatial patterns in ecology. An open problem in the biological sciences is characterising spatio-temporal data and understanding how changes at the local scale…

Topological Data Analysis