paper-with-me

홈 › Papers

RED: Reinforced Encoder-Decoder Networks for Action Anticipation

2017-07-16 · Jiyang Gao, Zhenheng Yang, Ram Nevatia

Action anticipation aims to detect an action before it happens. Many real world applications in robotics and surveillance are related to this predictive capability. Current methods address this problem by first anticipating visual representations of future frames and then categorizing the anticipated representations to actions. However, anticipation is based on a single past frame's representation, which ignores the history trend. Besides, it can only anticipate a fixed future time. We propose a Reinforced Encoder-Decoder (RED) network for action anticipation. RED takes multiple history representations as input and learns to anticipate a sequence of future representations. One salient aspect of RED is that a reinforcement module is adopted to provide sequence-level supervision; the reward function is designed to encourage the system to make correct predictions as early as possible. We test RED on TVSeries, THUMOS-14 and TV-Human-Interaction datasets for action anticipation and achieve state-of-the-art performance on all datasets.

📄 PDF Abstract BibTeX arXiv:1707.04818

Code (1)

rajskar/CS763Project pytorch

Tasks

Action AnticipationDecoder

Similar Papers 제목 키워드 기반

DRIVE: Deep Reinforced Accident Anticipation with Visual Explanation

2021-07-21 · ICCV 2021 10 · Wentao Bao, Qi Yu, Yu Kong

Traffic accident anticipation aims to accurately and promptly predict the occurrence of a future accident from dashcam videos, which is vital for a safety-guaranteed self-driving system. To encourage an early and accurat…

Accident AnticipationDecision Making

Bidirectional Action Sequence Learning for Long-term Action Anticipation with Large Language Models

2025-08-01 · Yuji Sato, Yasunori Ishii, Takayoshi Yamashita arxiv

Video-based long-term action anticipation is crucial for early risk detection in areas such as automated driving and robotics. Conventional approaches extract features from past actions using encoders and predict future …

Action Anticipation

QueryMamba: A Mamba-Based Encoder-Decoder Architecture with a Statistical Verb-Noun Interaction Module for Video Action Forecasting @ Ego4D Long-Term Action Anticipation Challenge 2024

2024-07-04 · Zeyun Zhong, Manuel Martin, Frederik Diederichs, Juergen Beyerer

This report presents a novel Mamba-based encoder-decoder architecture, QueryMamba, featuring an integrated verb-noun interaction module that utilizes a statistical verb-noun co-occurrence matrix to enhance video action f…

Action AnticipationDecoderLong Term Action AnticipationMamba

Temporal Context Consistency Above All: Enhancing Long-Term Anticipation by Learning and Enforcing Temporal Constraints

2024-12-27 · Alberto Maté, Mariella Dimiccoli

This paper proposes a method for long-term action anticipation (LTA), the task of predicting action labels and their duration in a video given the observation of an initial untrimmed video interval. We build on an encode…

Action AnticipationAction SegmentationAllDecoder+2

Technical Report for Ego4D Long Term Action Anticipation Challenge 2023

2023-07-04 · Tatsuya Ishibashi, Kosuke Ono, Noriyuki Kugo, Yuji Sato

In this report, we describe the technical details of our approach for the Ego4D Long-Term Action Anticipation Challenge 2023. The aim of this task is to predict a sequence of future actions that will take place at an arb…

Action AnticipationDecoderLong Term Action Anticipation