paper-with-me

Papers

Motion Aware Self-Supervision for Generic Event Boundary Detection

2022-10-11 · Ayush K. Rai, Tarun Krishna, Julia Dietlmeier, Kevin McGuinness, Alan F. Smeaton, Noel E. O'Connor

The task of Generic Event Boundary Detection (GEBD) aims to detect moments in videos that are naturally perceived by humans as generic and taxonomy-free event boundaries. Modeling the dynamically evolving temporal and spatial changes in a video makes GEBD a difficult problem to solve. Existing approaches involve very complex and sophisticated pipelines in terms of architectural design choices, hence creating a need for more straightforward and simplified approaches. In this work, we address this issue by revisiting a simple and effective self-supervised method and augment it with a differentiable motion feature learning module to tackle the spatial and temporal diversities in the GEBD task. We perform extensive experiments on the challenging Kinetics-GEBD and TAPOS datasets to demonstrate the efficacy of the proposed approach compared to the other self-supervised state-of-the-art methods. We also show that this simple self-supervised approach learns motion features without any explicit motion-specific pretext task.

📄 PDF Abstract BibTeX arXiv:2210.05574

Code (1)

rayush7/motion_ssl_gebd 공식 구현

Tasks

Boundary DetectionGeneric Event Boundary Detection

Similar Papers 제목 키워드 기반

EventDrive: Event Cameras for Vision-Language Driving Intelligence

2026-06-16 · Dongyue Lu, Rong Li, Ao Liang, Lingdong Kong 외 arxiv

Event cameras sense the world through asynchronous brightness changes with microsecond latency and high dynamic range, offering motion fidelity far beyond frame-based sensors and capturing temporal structure that convent…

Trajectory ForecastingAutonomous Driving

Does Visual Self-Supervision Improve Learning of Speech Representations for Emotion Recognition?

2020-05-04 · Abhinav Shukla, Stavros Petridis, Maja Pantic

Self-supervised learning has attracted plenty of recent research interest. However, most works for self-supervision in speech are typically unimodal and there has been limited work that studies the interaction between au…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Emotion RecognitionFace Reconstruction+5

Domain-aware Self-supervised Pre-training for Weakly-supervised Meme Analysis

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Existing self-supervised learning strategies are constrained to a limited set of trivial and generic downstream tasks that predominantly target uni-modal applications. This has isolated progress for imperative multi-moda…

Representation LearningSelf-Supervised Learning

Match-Any-Events: Zero-Shot Motion-Robust Feature Matching Across Wide Baselines for Event Cameras

2026-04-20 · Ruijun Zhang, Hang Su, Kostas Daniilidis, Ziyun Wang arxiv

Event cameras have recently shown promising capabilities in instantaneous motion estimation due to their robustness to low light and fast motions. However, computing wide-baseline correspondence between two arbitrary vie…

Motion Synthesis

Motion Deblurring with Real Events

2021-09-28 · ICCV 2021 10 · Fang Xu, Lei Yu, Bishan Wang, Wen Yang 외

In this paper, we propose an end-to-end learning framework for event-based motion deblurring in a self-supervised manner, where real-world events are exploited to alleviate the performance degradation caused by data inco…

Deblurring