Long-video Activity Recognition
1개 벤치마크 · 논문 6편 · 이 태스크의 논문 보기 →
Benchmarks
Breakfast
Most implemented
Timeception for Complex Action Recognition
Papers
ST(OR)2: Spatio-Temporal Object Level Reasoning for Activity Recognition in the Operating Room
Surgical robotics holds much promise for improving patient safety and clinician experience in the Operating Room (OR). However, it also comes with new challenges, requiring strong team coordination and effective OR manag…
Action ClassificationActivity RecognitionLong-video Activity RecognitionManagement+1Towards Weakly Supervised End-to-end Learning for Long-video Action Recognition
Developing end-to-end action recognition models on long videos is fundamental and crucial for long-video action understanding. Due to the unaffordable cost of end-to-end training on the whole long videos, existing works …
Action ClassificationAction RecognitionAction SegmentationAction Understanding+5Graph-Based High-Order Relation Modeling for Long-Term Action Recognition
Long-term actions involve many important visual concepts, e.g., objects, motions, and sub-actions, and there are various relations among these concepts, which we call basic relations. These basic relations will joint…
Action RecognitionLong-video Activity RecognitionRelationVideo Classification+1VideoGraph: Recognizing Minutes-Long Human Activities in Videos
Many human activities take minutes to unfold. To represent them, related works opt for statistical pooling, which neglects the temporal structure. Others opt for convolutional methods, as CNN and Non-Local. While success…
Long-video Activity RecognitionVideo ClassificationTimeception for Complex Action Recognition
This paper focuses on the temporal aspect for recognizing human activities in videos; an important visual cue that has long been undervalued. We revisit the conventional definition of activity and restrict it to Complex …
Action ClassificationAction RecognitionLong-video Activity RecognitionVideo ClassificationActionVLAD: Learning spatio-temporal aggregation for action classification
In this work, we introduce a new video representation for action classification that aggregates local convolutional features across the entire spatio-temporal extent of the video. We do so by integrating state-of-the-art…
Action ClassificationClassificationGeneral ClassificationLong-video Activity Recognition+1