paper-with-me

Papers

PKU-MMD: A Large Scale Benchmark for Continuous Multi-Modal Human Action Understanding

2017-03-22 · Chunhui Liu, Yueyu Hu, Yanghao Li, Sijie Song, Jiaying Liu

Despite the fact that many 3D human activity benchmarks being proposed, most existing action datasets focus on the action recognition tasks for the segmented videos. There is a lack of standard large-scale benchmarks, especially for current popular data-hungry deep learning based methods. In this paper, we introduce a new large scale benchmark (PKU-MMD) for continuous multi-modality 3D human action understanding and cover a wide range of complex human activities with well annotated information. PKU-MMD contains 1076 long video sequences in 51 action categories, performed by 66 subjects in three camera views. It contains almost 20,000 action instances and 5.4 million frames in total. Our dataset also provides multi-modality data sources, including RGB, depth, Infrared Radiation and Skeleton. With different modalities, we conduct extensive experiments on our dataset in terms of two scenarios and evaluate different methods by various metrics, including a new proposed evaluation protocol 2D-AP. We believe this large-scale dataset will benefit future researches on action detection for the community.

📄 PDF Abstract BibTeX arXiv:1703.07475

Code (0)

등록된 구현이 없습니다.

Tasks

Action DetectionAction RecognitionAction UnderstandingTemporal Action Localization

Similar Papers 제목 키워드 기반

BEACON: A Multimodal Dataset for Learning Behavioral Fingerprints from Gameplay Data

2026-05-11 · Ishpuneet Singh, Gursmeep Kaur, Uday Pratap Singh Atwal, Guramrit Singh 외 arxiv

Continuous authentication in high-stakes digital environments requires datasets with fine-grained behavioral signals under realistic cognitive and motor demands. But current benchmarks are often limited by small scale, u…

Representation Learning

Generating Large-scale Dynamic Optimization Problem Instances Using the Generalized Moving Peaks Benchmark

2021-07-23 · Mohammad Nabi Omidvar, Danial Yazdani, Juergen Branke, XiaoDong Li 외

This document describes the generalized moving peaks benchmark (GMPB) and how it can be used to generate problem instances for continuous large-scale dynamic optimization problems. It presents a set of 15 benchmark probl…

OrthoTrack: Continuous 6-DoF UAV Trajectory Estimation Anchored in Public Orthophotos

2026-06-24 · Oussema Dhaouadi, Zuria Bauer, Johannes Michael Meier, Olaf Wysocki 외 arxiv

Continuous 6-DoF pose estimation is essential for autonomous UAV operations. Yet, existing visual odometry and SLAM methods accumulate drift and yield only relative, up-to-scale trajectories. Single-frame geo-localizatio…

Pose EstimationVisual Odometry

Multimodal Clinical Benchmark for Emergency Care (MC-BEC): A Comprehensive Benchmark for Evaluating Foundation Models in Emergency Medicine

2023-11-07 · NeurIPS 2023 11

We propose the Multimodal Clinical Benchmark for Emergency Care (MC-BEC), a comprehensive benchmark for evaluating foundation models in Emergency Medicine using a dataset of 100K+ continuously monitored Emergency Departm…

Decompensation

Efficient Large-Scale Multi-Modal Classification

2018-02-06 · D. Kiela, E. Grave, A. Joulin, T. Mikolov

While the incipient internet was largely text-based, the modern digital world is becoming increasingly multi-modal. Here, we examine multi-modal classification where one modality is discrete, e.g. text, and the other is …

ClassificationComputational EfficiencyGeneral ClassificationMulti-modal Classification