paper-with-me

홈 › Papers

TimeGate: Conditional Gating of Segments in Long-range Activities

2020-04-03 · Noureldien Hussein, Mihir Jain, Babak Ehteshami Bejnordi

When recognizing a long-range activity, exploring the entire video is exhaustive and computationally expensive, as it can span up to a few minutes. Thus, it is of great importance to sample only the salient parts of the video. We propose TimeGate, along with a novel conditional gating module, for sampling the most representative segments from the long-range activity. TimeGate has two novelties that address the shortcomings of previous sampling methods, as SCSampler. First, it enables a differentiable sampling of segments. Thus, TimeGate can be fitted with modern CNNs and trained end-to-end as a single and unified model.Second, the sampling is conditioned on both the segments and their context. Consequently, TimeGate is better suited for long-range activities, where the importance of a segment heavily depends on the video context.TimeGate reduces the computation of existing CNNs on three benchmarks for long-range activities: Charades, Breakfast and MultiThumos. In particular, TimeGate reduces the computation of I3D by 50% while maintaining the classification accuracy.

📄 PDF Abstract BibTeX arXiv:2004.01808

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TIMEGATE: Sustainable Time-Boxed Promotion Gates for Continual ML Adaptation Under Resource Constraints

2026-05-27 · Abhijit Chakraborty, Suddhasvatta Das, Yash Shah, Vivek Gupta 외 arxiv

As machine learning(ML) systems evolve to continual adaptation, each re-training cycle uses compute, annotation, and energy. We introduce TIMEGATE, a policy layer managing adaptation by budgeting time, labeling, training…

Unifying Multitrack Music Arrangement via Reconstruction Fine-Tuning and Efficient Tokenization

2024-08-27 · Longshen Ou, Jingwei Zhao, Ziyu Wang, Gus Xia 외

Automatic music arrangement streamlines the creation of musical variants for composers and arrangers, reducing reliance on extensive music expertise. However, existing methods suffer from inefficient tokenization, underu…

Language ModelingLanguage ModellingMusic Generation

Revisiting Kernel Temporal Segmentation as an Adaptive Tokenizer for Long-form Video Understanding

2023-09-20 · Mohamed Afham, Satya Narayan Shukla, Omid Poursaeed, Pengchuan Zhang 외

While most modern video understanding models operate on short-range clips, real-world videos are often several minutes long with semantically consistent segments of variable length. A common approach to process long vide…

Action LocalizationFormTemporal Action LocalizationVideo Classification+1

MoE-Gyro: Self-Supervised Over-Range Reconstruction and Denoising for MEMS Gyroscopes

2025-05-27 · Feiyang Pan, Shenghe Zheng, Chunyan Yin, Guangbin Dou

MEMS gyroscopes play a critical role in inertial navigation and motion control applications but typically suffer from a fundamental trade-off between measurement range and noise performance. Existing hardware-based solut…

BenchmarkingDenoisingMixture-of-Experts

AdmTree: Compressing Lengthy Context with Adaptive Semantic Trees

2025-12-04 · Yangning Li, Shaoshen Chen, Yinghui Li, Yankai Chen 외 arxiv

The quadratic complexity of self-attention constrains Large Language Models (LLMs) in processing long contexts, a capability essential for many advanced applications. Context compression aims to alleviate this computatio…