paper-with-me

홈 › Papers

Home Action Genome: Cooperative Compositional Action Understanding

2021-05-11 · CVPR 2021 1 · Nishant Rai, Haofeng Chen, Jingwei Ji, Rishi Desai, Kazuki Kozuka, Shun Ishizaka, Ehsan Adeli, Juan Carlos Niebles

Existing research on action recognition treats activities as monolithic events occurring in videos. Recently, the benefits of formulating actions as a combination of atomic-actions have shown promise in improving action understanding with the emergence of datasets containing such annotations, allowing us to learn representations capturing this information. However, there remains a lack of studies that extend action composition and leverage multiple viewpoints and multiple modalities of data for representation learning. To promote research in this direction, we introduce Home Action Genome (HOMAGE): a multi-view action dataset with multiple modalities and view-points supplemented with hierarchical activity and atomic action labels together with dense scene composition labels. Leveraging rich multi-modal and multi-view settings, we propose Cooperative Compositional Action Understanding (CCAU), a cooperative learning framework for hierarchical action recognition that is aware of compositional action elements. CCAU shows consistent performance improvements across all modalities. Furthermore, we demonstrate the utility of co-learning compositions in few-shot action recognition by achieving 28.6% mAP with just a single sample.

📄 PDF Abstract BibTeX arXiv:2105.05226

Code (1)

nishantrai18/homage 공식 구현 pytorch

Tasks

Action RecognitionAction UnderstandingFew-Shot action recognitionFew Shot Action RecognitionRepresentation LearningVideo Classification

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

AGQA 2.0: An Updated Benchmark for Compositional Spatio-Temporal Reasoning

2022-04-12 · Madeleine Grunde-McLaughlin, Ranjay Krishna, Maneesh Agrawala

Prior benchmarks have analyzed models' answers to questions about videos in order to measure visual compositional reasoning. Action Genome Question Answering (AGQA) is one such benchmark. AGQA provides a training/test sp…

Question Answering

Revisiting spatio-temporal layouts for compositional action recognition

2021-11-02 · Gorjan Radevski, Marie-Francine Moens, Tinne Tuytelaars

Recognizing human actions is fundamentally a spatio-temporal reasoning problem, and should be, at least to some extent, invariant to the appearance of the human and the objects involved. Motivated by this hypothesis, in …

Action ClassificationAction DetectionAction RecognitionFew-Shot action recognition+3

LEMMA: A Multi-view Dataset for Learning Multi-agent Multi-task Activities

2020-07-31 · ECCV 2020 8 · Baoxiong Jia, Yixin Chen, Siyuan Huang, Yixin Zhu 외

Understanding and interpreting human actions is a long-standing challenge and a critical indicator of perception in artificial intelligence. However, a few imperative components of daily human activities are largely miss…

Action RecognitionAction UnderstandingHuman-Object Interaction DetectionLEMMA+1

Inferring Past Human Actions in Homes with Abductive Reasoning

2022-10-24 · Clement Tan, Chai Kiat Yeo, Cheston Tan, Basura Fernando

Abductive reasoning aims to make the most likely inference for a given set of incomplete observations. In this paper, we introduce "Abductive Past Action Inference", a novel research task aimed at identifying the past ac…

DecoderGraph Neural Network

gghic: A Versatile R Package for Exploring and Visualizing 3D Genome Organization

2024-12-04 · Minghao Jiang, Duohui Jing, Jason W. H. Wong

Motivation: The three-dimensional (3D) organization of the genome plays a critical role in regulating gene expression and maintaining cellular homeostasis. Disruptions in this spatial organization can result in abnormal …