paper-with-me

Papers

Multi-View Region Adaptive Multi-temporal DMM and RGB Action Recognition

2019-04-12 · Mahmoud Al-Faris, John P. Chiverton, Yanyan Yang, David L. Ndzi

Human action recognition remains an important yet challenging task. This work proposes a novel action recognition system. It uses a novel Multiple View Region Adaptive Multi-resolution in time Depth Motion Map (MV-RAMDMM) formulation combined with appearance information. Multiple stream 3D Convolutional Neural Networks (CNNs) are trained on the different views and time resolutions of the region adaptive Depth Motion Maps. Multiple views are synthesised to enhance the view invariance. The region adaptive weights, based on localised motion, accentuate and differentiate parts of actions possessing faster motion. Dedicated 3D CNN streams for multi-time resolution appearance information (RGB) are also included. These help to identify and differentiate between small object interactions. A pre-trained 3D-CNN is used here with fine-tuning for each stream along with multiple class Support Vector Machines (SVM)s. Average score fusion is used on the output. The developed approach is capable of recognising both human action and human-object interaction. Three public domain datasets including: MSR 3D Action,Northwestern UCLA multi-view actions and MSR 3D daily activity are used to evaluate the proposed solution. The experimental results demonstrate the robustness of this approach compared with state-of-the-art algorithms.

📄 PDF Abstract BibTeX arXiv:1904.06074

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionHuman-Object Interaction DetectionTemporal Action Localization

Similar Papers 제목 키워드 기반

SMA-Hyper: Spatiotemporal Multi-View Fusion Hypergraph Learning for Traffic Accident Prediction

2024-07-24 · Xiaowei Gao, James Haworth, Ilya Ilyankou, Xianghui Zhang 외

Predicting traffic accidents is the key to sustainable city management, which requires effective address of the dynamic and complex spatiotemporal characteristics of cities. Current data-driven models often struggle with…

Contrastive LearningGraph LearningManagement

Spatial-Temporal Graph Learning with Adversarial Contrastive Adaptation

2023-06-19 · Qianru Zhang, Chao Huang, Lianghao Xia, Zheng Wang 외

Spatial-temporal graph learning has emerged as a promising solution for modeling structured spatial-temporal data and learning region representations for various urban sensing tasks such as crime forecasting and traffic …

Contrastive LearningGraph LearningSelf-Supervised Learning

ACT-R: Adaptive Camera Trajectories for Single View 3D Reconstruction

2025-05-13 · Yizhi Wang, Mingrui Zhao, Ali Mahdavi-Amiri, Hao Zhang

We introduce the simple idea of adaptive view planning to multi-view synthesis, aiming to improve both occlusion revelation and 3D consistency for single-view 3D reconstruction. Instead of producing an unordered set of v…

3D ReconstructionMulti-View 3D ReconstructionSingle-View 3D Reconstruction

FrustumFormer: Adaptive Instance-aware Resampling for Multi-view 3D Detection

2023-01-10 · CVPR 2023 1 · Yuqi Wang, Yuntao Chen, Zhaoxiang Zhang

The transformation of features from 2D perspective space to 3D space is essential to multi-view 3D object detection. Recent approaches mainly focus on the design of view transformation, either pixel-wisely lifting perspe…

3D Object Detectionobject-detectionObject Detection

Temporally Aware Densification for Dynamic 3D Gaussian Splatting

2026-06-22 · Vikram Sandu, Mayurdeep Pathak, Rajiv Soundararajan arxiv

Despite modeling temporal motion, dynamic 3D Gaussian Splatting (3DGS) methods still inherit a static densification strategy that is ill-suited for dynamic scenes. This neglect of temporal behavior leads to under-reconst…

Dynamic Reconstruction