paper-with-me

Papers

Multi-modal Egocentric Activity Recognition using Audio-Visual Features

2018-07-02 · Mehmet Ali Arabaci, Fatih Özkan, Elif Surer, Peter Jančovič, Alptekin Temizel

Egocentric activity recognition in first-person videos has an increasing importance with a variety of applications such as lifelogging, summarization, assisted-living and activity tracking. Existing methods for this task are based on interpretation of various sensor information using pre-determined weights for each feature. In this work, we propose a new framework for egocentric activity recognition problem based on combining audio-visual features with multi-kernel learning (MKL) and multi-kernel boosting (MKBoost). For that purpose, firstly grid optical-flow, virtual-inertia feature, log-covariance, cuboid are extracted from the video. The audio signal is characterized using a "supervector", obtained based on Gaussian mixture modelling of frame-level features, followed by a maximum a-posteriori adaptation. Then, the extracted multi-modal features are adaptively fused by MKL classifiers in which both the feature and kernel selection/weighing and recognition tasks are performed together. The proposed framework was evaluated on a number of egocentric datasets. The results showed that using multi-modal features with MKL outperforms the existing methods.

📄 PDF Abstract BibTeX arXiv:1807.00612

Code (0)

등록된 구현이 없습니다.

Tasks

Activity RecognitionEgocentric Activity RecognitionOptical Flow Estimation

Similar Papers 제목 키워드 기반

Towards Continual Egocentric Activity Recognition: A Multi-modal Egocentric Activity Dataset for Continual Learning

2023-01-26 · Linfeng Xu, Qingbo Wu, Lili Pan, Fanman Meng 외

With the rapid development of wearable cameras, a massive collection of egocentric video for first-person visual perception becomes available. Using egocentric videos to predict first-person activity faces many challenge…

Activity RecognitionContinual LearningEgocentric Activity RecognitionHuman Activity Recognition

Egocentric Activity Recognition with Multimodal Fisher Vector

2016-01-25 · Sibo Song, Ngai-Man Cheung, Vijay Chandrasekhar, Bappaditya Mandal 외

With the increasing availability of wearable devices, research on egocentric activity recognition has received much attention recently. In this paper, we build a Multimodal Egocentric Activity dataset which includes egoc…

Activity RecognitionEgocentric Activity Recognition

EPIC-Fusion: Audio-Visual Temporal Binding for Egocentric Action Recognition

2019-08-22 · ICCV 2019 10 · Evangelos Kazakos, Arsha Nagrani, Andrew Zisserman, Dima Damen

We focus on multi-modal fusion for egocentric action recognition, and propose a novel architecture for multi-modal temporal-binding, i.e. the combination of modalities within a range of temporal offsets. We train the arc…

Action RecognitionEgocentric Activity Recognition

MAND: Modality-Aware Novelty Detection for Open-World Egocentric Activity Recognition

2026-03-17 · Hyejeong Im, Wonseon Lim, Dae-Won Kim arxiv

Multimodal egocentric activity recognition integrates visual and inertial cues for robust first-person behavior understanding. However, deploying such systems in open-world environments requires detecting novel activitie…

Egocentric Activity RecognitionContinual LearningActivity Detection

With a Little Help from my Temporal Context: Multimodal Egocentric Action Recognition

2021-11-01 · Evangelos Kazakos, Jaesung Huh, Arsha Nagrani, Andrew Zisserman 외

In egocentric videos, actions occur in quick succession. We capitalise on the action's temporal context and propose a method that learns to attend to surrounding actions in order to improve recognition performance. To in…

Action RecognitionLanguage ModelingLanguage Modelling