paper-with-me

Papers Spatio-temporal Action Recognition

“Spatio-temporal Action Recognition” 태그가 달린 논문 9편 · 필터 해제

DVFL-Net: A Lightweight Distilled Video Focal Modulation Network for Spatio-Temporal Action Recognition

2025-07-16 · Hayat Ullah, Muhammad Ali Shafique, Abbas Khan, Arslan Munir

The landscape of video recognition has evolved significantly, shifting from traditional Convolutional Neural Networks (CNNs) to Transformer-based architectures for improved accuracy. While 3D CNNs have been effective at …

BenchmarkingKnowledge DistillationSpatio-temporal Action RecognitionTemporal Action Localization+2

O-TALC: Steps Towards Combating Oversegmentation within Online Action Segmentation

2024-04-10 · Matthew Kent Myers, Nick Wright, A. Stephen McGough, Nicholas Martin

Online temporal action segmentation shows a strong potential to facilitate many HRI tasks where extended human action sequences must be tracked and understood in real time. Traditional action segmentation approaches, how…

Action RecognitionAction SegmentationSegmentationSpatio-temporal Action Recognition+1

Point3D: tracking actions as moving points with 3D CNNs

2022-03-20 · Shentong Mo, Jingfei Xia, Xiaoqing Tan, Bhiksha Raj

Spatio-temporal action recognition has been a challenging task that involves detecting where and when actions occur. Current state-of-the-art action detectors are mostly anchor-based, requiring sensitive anchor designs a…

Action ClassificationAction LocalizationAction RecognitionSpatio-temporal Action Recognition

A Study On the Effects of Pre-processing On Spatio-temporal Action Recognition Using Spiking Neural Networks Trained with STDP

2021-05-31 · El-Assal Mireille, Tirilly Pierre, Bilasco Ioan Marius

There has been an increasing interest in spiking neural networks in recent years. SNNs are seen as hypothetical solutions for the bottlenecks of ANNs in pattern recognition, such as energy efficiency. But current methods…

Action RecognitionSpatio-temporal Action RecognitionVideo ClassificationVideo Understanding

Toward Accurate Person-level Action Recognition in Videos of Crowded Scenes

2020-10-16 · Li Yuan, Yichen Zhou, Shuning Chang, Ziyuan Huang 외

Detecting and recognizing human action in videos with crowded scenes is a challenging problem due to the complex environment and diversity events. Prior works always fail to deal with this problem in two aspects: (1) lac…

Action RecognitionAction Recognition In VideosDiversitySemantic Segmentation+1

Uncertainty-Aware Weakly Supervised Action Detection from Untrimmed Videos

2020-07-21 · ECCV 2020 8 · Anurag Arnab, Chen Sun, Arsha Nagrani, Cordelia Schmid

Despite the recent advances in video classification, progress in spatio-temporal action recognition has lagged behind. A major contributing factor has been the prohibitive cost of annotating videos frame-by-frame. In thi…

Action DetectionAction RecognitionMultiple Instance LearningSpatio-temporal Action Recognition+1

Recognizing Manipulation Actions from State-Transformations

2019-06-12 · Nachwa Aboubakr, James L. Crowley, Remi Ronfard

Manipulation actions transform objects from an initial state into a final state. In this paper, we report on the use of object state transitions as a mean for recognizing manipulation actions. Our method is inspired by t…

Action RecognitionObjectSpatio-temporal Action Recognition

Spatio-temporal Action Recognition: A Survey

2019-01-27 · Amlaan Bhoi

The task of action recognition or action detection involves analyzing videos and determining what action or motion is being performed. The primary subject of these videos are predominantly humans performing some action. …

Action DetectionAction LocalizationAction RecognitionSpatio-temporal Action Recognition+3

AVA: A Video Dataset of Spatio-temporally Localized Atomic Visual Actions

2017-05-23 · CVPR 2018 6 · Chunhui Gu, Chen Sun, David A. Ross, Carl Vondrick 외

This paper introduces a video dataset of spatio-temporally localized Atomic Visual Actions (AVA). The AVA dataset densely annotates 80 atomic visual actions in 430 15-minute video clips, where actions are localized in sp…

Actin DetectionAction DetectionAction LocalizationAction Recognition+3
1–9 / 9