paper-with-me

Papers

AE-Net:Adjoint Enhancement Network for Efficient Action Recognition in Video Understanding

2022-07-21 · TMM 2022 7 · Bin Wang, Chunsheng Liu, Faliang Chang, Wenqian Wang and Nanjun Li

Action recognition in video understanding is a challenging task, largely because of the complexity and difficulty in temporal modeling, making it suffer from motion information loss and misalignment of temporal attention in spatial dimensions. To overcome these difficulties, we propose a novel temporal modeling method called Adjoint Enhancement Network (AE-Net), which can fully explore clues of motion and time in the long-range structure. The AE-Net mainly consists of two new modules: the Initial Adjoint Enhancement Module (IAE-Module), which deals with shallow features; and the Global Adjoint Enhancement Module (GAE-Module), which deals with global features. With a novel mechanism of parallel spatio-temporal convolution and difference fusion, the IAE-Module is to enhance the degree of motion transformation in shallow network features, exciting the potential of motion flow and avoiding motion information loss. The GAE-Module is proposed to improve the local temporal representation in long-range structures by feeding the enhanced feature differences into a spatial cascade module with residuals to resolve the misalignment of temporal attention in the spatial dimension.The experimental results show that our AE-Net can achieve state-of-the-art results in Something-Something V1, UCF101 and HMDB-51 datasets.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionVideo Understanding

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

ARID: A New Dataset for Recognizing Action in the Dark

2020-06-06 · Yuecong Xu, Jianfei Yang, Haozhi Cao, Kezhi Mao 외

The task of action recognition in dark videos is useful in various scenarios, e.g., night surveillance and self-driving at night. Though progress has been made in the action recognition task for videos in normal illumina…

Action Recognition

Cross-Enhancement Transform Two-Stream 3D ConvNets for Action Recognition

2019-08-19 · Dong Cao, Lisha Xu, Dong-dong Zhang

Action recognition is an important research topic in computer vision. It is the basic work for visual understanding and has been applied in many fields. Since human actions can vary in different environments, it is diffi…

Action RecognitionAutonomous DrivingAutonomous VehiclesOptical Flow Estimation+2

The Role of Video Generation in Enhancing Data-Limited Action Understanding

2025-05-26 · Wei Li, Dezhao Luo, Dongbao Yang, Zhenhang Li 외

Video action understanding tasks in real-world scenarios always suffer data limitations. In this paper, we address the data-limited action understanding problem by bridging data scarcity. We propose a novel method that e…

Action RecognitionAction UnderstandingVideo GenerationZero-Shot Action Recognition

IndGIC: Supervised Action Recognition under Low Illumination

2023-08-29 · Jingbo Zeng

Technologies of human action recognition in the dark are gaining more and more attention as huge demand in surveillance, motion control and human-computer interaction. However, because of limitation in image enhancement …

Action RecognitionImage EnhancementTemporal Action Localization

Learnable Sampling 3D Convolution for Video Enhancement and Action Recognition

2020-11-22 · Shuyang Gu, Jianmin Bao, Dong Chen

A key challenge in video enhancement and action recognition is to fuse useful information from neighboring frames. Recent works suggest establishing accurate correspondences between neighboring frames before fusing tempo…

Action RecognitionDenoisingSuper-ResolutionVideo Denoising+2