paper-with-me

Papers

Latency Matters: Real-Time Action Forecasting Transformer

2023-01-01 · CVPR 2023 1 · Harshayu Girase, Nakul Agarwal, Chiho Choi, Karttikeya Mangalam

We present RAFTformer, a real-time action forecasting transformer for latency aware real-world action forecasting applications. RAFTformer is a two-stage fully transformer based architecture which consists of a video transformer backbone that operates on high resolution, short range clips and a head transformer encoder that temporally aggregates information from multiple short range clips to span a long-term horizon. Additionally, we propose a self-supervised shuffled causal masking scheme to improve model generalization during training. Finally, we also propose a real-time evaluation setting that directly couples model inference latency to overall forecasting performance and brings forth an hitherto overlooked trade-off between latency and action forecasting performance. Our parsimonious network design facilitates RAFTformer inference latency to be 9x smaller than prior works at the same forecasting accuracy. Owing to its two-staged design, RAFTformer uses 94% less training compute and 90% lesser training parameters to outperform prior state-of-the-art baselines by 4.9 points on EGTEA Gaze+ and by 1.4 points on EPIC-Kitchens-100 dataset, as measured by Top-5 recall (T5R) in the offline setting. In the real-time setting, RAFTformer outperforms prior works by an even greater margin of upto 4.4 T5R points on the EPIC-Kitchens-100 dataset. Project Webpage: https://karttikeya.github.io/publication/RAFTformer/

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

Enhancing 5G O-RAN Communication Efficiency Through AI-Based Latency Forecasting

2025-02-25 · Raúl Parada, Ebrahim Abu-Helalah, Jordi Serra, Anton Aguilar 외

The increasing complexity and dynamic nature of 5G open radio access networks (O-RAN) pose significant challenges to maintaining low latency, high throughput, and resource efficiency. While existing methods leverage mach…

Management

Make a Video Call with LLM: A Measurement Campaign over Six Mainstream Apps

2025-10-01 · Jiayang Xu, Xiangjie Huang, Zijie Li, Antariksh Verma 외 arxiv

In 2025, Large Language Model (LLM) services have launched a new feature -- AI video chat -- allowing users to interact with AI agents via real-time video communication (RTC), just like chatting with real people. Despite…

ReLaMix: Residual Latency-Aware Mixing for Delay-Robust Financial Time-Series Forecasting

2026-03-21 · Tianyou Lai, Wentao Yue, Jiayi Zhou, Chaoyuan Hao 외 arxiv

Financial time-series forecasting in real-world high-frequency markets is often hindered by delayed or partially stale observations caused by asynchronous data acquisition and transmission latency. To better reflect such…

FUTURE-VLA: Forecasting Unified Trajectories Under Real-time Execution

2026-02-05 · Jingjing Fan, Yushan Liu, Shoujie Li, Botao Ren 외 arxiv

General vision-language models increasingly support unified spatiotemporal reasoning over long video streams, yet deploying such capabilities on robots remains constrained by the prohibitive latency of processing long-ho…

Uniqueness Bias: Why It Matters, How to Curb It

2024-08-13 · Bent Flyvbjerg, Alexander Budzier, M. D. Christodoulou, M. Zottoli

The paper explores "uniqueness bias," a behavioral bias defined as the tendency of planners and managers to see their decisions as singular. For the first time, uniqueness bias is correlated with forecasting accuracy and…