paper-with-me

홈 › Papers

ReActNet: Temporal Localization of Repetitive Activities in Real-World Videos

2019-10-14 · Giorgos Karvounas, Iason Oikonomidis, Antonis Argyros

We address the problem of temporal localization of repetitive activities in a video, i.e., the problem of identifying all segments of a video that contain some sort of repetitive or periodic motion. To do so, the proposed method represents a video by the matrix of pairwise frame distances. These distances are computed on frame representations obtained with a convolutional neural network. On top of this representation, we design, implement and evaluate ReActNet, a lightweight convolutional neural network that classifies a given frame as belonging (or not) to a repetitive video segment. An important property of the employed representation is that it can handle repetitive segments of arbitrary number and duration. Furthermore, the proposed training process requires a relatively small number of annotated videos. Our method raises several of the limiting assumptions of existing approaches regarding the contents of the video and the types of the observed repetitive activities. Experimental results on recent, publicly available datasets validate our design choices, verify the generalization potential of ReActNet and demonstrate its superior performance in comparison to the current state of the art.

📄 PDF Abstract BibTeX arXiv:1910.06096

Code (0)

등록된 구현이 없습니다.

Tasks

Temporal Localization

Similar Papers 제목 키워드 기반

TransRAC: Encoding Multi-scale Temporal Correlation with Transformers for Repetitive Action Counting

2022-04-03 · CVPR 2022 1 · Huazhang Hu, Sixun Dong, Yiqun Zhao, Dongze Lian 외

Counting repetitive actions are widely seen in human activities such as physical exercise. Existing methods focus on performing repetitive action counting in short videos, which is tough for dealing with longer videos in…

Repetitive Action Counting

Advanced Gesture Recognition in Autism: Integrating YOLOv7, Video Augmentation and VideoMAE for Video Analysis

2024-10-12 · Amit Kumar Singh, Trapti Shrivastava, Vrijendra Singh

Deep learning and advancements in contactless sensors have significantly enhanced our ability to understand complex human activities in healthcare settings. In particular, deep learning models utilizing computer vision h…

Gesture Recognition

MM-SEAL: A Large-scale Video Dataset of Multi-person Multi-grained Spatio-temporally Action Localization

2022-04-06 · Shimin Chen, Wei Li, Chen Chen, Jianyang Gu 외

In this paper, we introduce a novel large-scale video dataset dubbed MM-SEAL for multi-person multi-grained spatio-temporal action localization among human daily life. We are the first to propose a new benchmark for mult…

Action LocalizationAction RecognitionSpatio-Temporal Action LocalizationTemporal Action Localization+1

Localization-Aware Multi-Scale Representation Learning for Repetitive Action Counting

2025-01-13 · Sujia Wang, Xiangwei Shen, Yansong Tang, Xin Dong 외

Repetitive action counting (RAC) aims to estimate the number of class-agnostic action occurrences in a video without exemplars. Most current RAC methods rely on a raw frame-to-frame similarity representation for period p…

Repetitive Action CountingRepresentation Learning

Bench-RNR: Dataset for Benchmarking Repetitive and Non-repetitive Scanning LiDAR for Infrastructure-based Vehicle Localization

2025-09-19 · Runxin Zhao, Chunxiang Wang, Hanyang Zhuang, Ming Yang arxiv

Vehicle localization using roadside LiDARs can provide centimeter-level accuracy for cloud-controlled vehicles while simultaneously serving multiple vehicles, enhanc-ing safety and efficiency. While most existing studies…

Point Clouds