CLAD: A Complex and Long Activities Dataset with Rich Crowdsourced Annotations
This paper introduces a novel activity dataset which exhibits real-life and diverse scenarios of complex, temporally-extended human activities and actions. The dataset presents a set of videos of actors performing everyday activities in a natural and unscripted manner. The dataset was recorded using a static Kinect 2 sensor which is commonly used on many robotic platforms. The dataset comprises of RGB-D images, point cloud data, automatically generated skeleton tracks in addition to crowdsourced annotations. Furthermore, we also describe the methodology used to acquire annotations through crowdsourcing. Finally some activity recognition benchmarks are presented using current state-of-the-art techniques. We believe that this dataset is particularly suitable as a testbed for activity recognition research but it can also be applicable for other common tasks in robotics/computer vision research such as object detection and human skeleton tracking.
Code (0)
등록된 구현이 없습니다.
Tasks
Activity Recognitionobject-detectionObject DetectionSimilar Papers 제목 키워드 기반
A hybrid machine learning framework for clad characteristics prediction in metal additive manufacturing
During the past decade, metal additive manufacturing (MAM) has experienced significant developments and gained much attention due to its ability to fabricate complex parts, manufacture products with functionally graded m…
Hybrid Machine LearningUnleashing Foundation Vision Models: Adaptive Transfer for Diverse Data-Limited Scientific Domains
In the big data era, the computer vision field benefits from large-scale datasets such as LAION-2B, LAION-400M, and ImageNet-21K, Kinetics, on which popular models like the ViT and ConvNeXt series have been pre-trained, …
CLaDMoP: Learning Transferrable Models from Successful Clinical Trials via LLMs
Many existing models for clinical trial outcome prediction are optimized using task-specific loss functions on trial phase-specific data. While this scheme may boost prediction for common diseases and drugs, it can hinde…
Large Language Modelparameter-efficient fine-tuningPredictionICLAD: In-Context Learning with Comparison-Guidance for Audio Deepfake Detection
Audio deepfakes pose a significant security threat, yet current state-of-the-art (SOTA) detection systems do not generalize well to realistic in-the-wild deepfakes. We introduce a novel \textbf{I}n-\textbf{C}ontext \text…
Audio Deepfake DetectionClass Label-aware Graph Anomaly Detection
Unsupervised GAD methods assume the lack of anomaly labels, i.e., whether a node is anomalous or not. One common observation we made from previous unsupervised methods is that they not only assume the absence of such ano…
Anomaly DetectionGraph Anomaly DetectionNode Classification