paper-with-me

Papers

Self-supervised Motion Representation via Scattering Local Motion Cues

2020-08-01 · ECCV 2020 8 · Yuan Tian, Zhaohui Che, Wenbo Bao, Guangtao Zhai, Zhiyong Gao

Motion representation is key to many computer vision problems but has never been well studied in the literature. Existing works usually rely on the optical flow estimation to assist other tasks such as action recognition, frame prediction, video segmentation, etc. In this paper, we leverage the massive unlabeled video data to learn an accurate explicit motion representation that aligns well with the semantic distribution of the moving objects. Our method subsumes a coarse-to-fine paradigm, which first decodes the low-resolution motion maps from the rich spatial-temporal features of the video, then adaptively upsamples the low-resolution maps to the full-resolution by considering the semantic cues. To achieve this, we propose a novel context guided motion upsampling layer that leverages the spatial context of video objects to learn the upsampling parameters in an efficient way. We prove the effectiveness of our proposed motion representation method on downstream video understanding tasks, e.g., action recognition task. Experimental results show that our method performs favorably against state-of-the-art methods.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionOptical Flow EstimationVideo SegmentationVideo Semantic SegmentationVideo Understanding

Similar Papers 제목 키워드 기반

Deep scattering network for speech emotion recognition

2021-05-11 · Premjeet Singh, Goutam Saha, Md Sahidullah

This paper introduces scattering transform for speech emotion recognition (SER). Scattering transform generates feature representations which remain stable to deformations and shifting in time and frequency without much …

Emotion RecognitionSpeech Emotion Recognition

Classification with Scattering Operators

2010-11-12 · Joan Bruna, Stéphane Mallat

A scattering vector is a local descriptor including multiscale and multi-direction co-occurrence information. It is computed with a cascade of wavelet decompositions and complex modulus. This scattering representation is…

ClassificationGeneral ClassificationHandwritten Digit RecognitionModel Selection+2

From Phase to Phenomenon: Self-Supervised Learning of Subsurface Scattering with Minimal Phase-shift Inputs

2026-06-28 · Arjun Majumdar, Raphael Braun, Andreas Engelhardt, Hendrik PA. Lensch arxiv

We propose a self-supervised pretraining framework for learning sub-surface scattering (SSS) light transport representations from minimal input. Our method leverages a stereo projector-camera setup that captures only eig…

Self-Supervised Learning

Scattering Networks for Hybrid Representation Learning

2018-09-17 · Edouard Oyallon, Sergey Zagoruyko, Gabriel Huang, Nikos Komodakis 외

Scattering networks are a class of designed Convolutional Neural Networks (CNNs) with fixed weights. We argue they can serve as generic representations for modelling images. In particular, by working in scattering space,…

Representation Learning

RedMotion: Motion Prediction via Redundancy Reduction

2023-06-19 · Royden Wagner, Omer Sahin Tas, Marvin Klemp, Carlos Fernandez 외

We introduce RedMotion, a transformer model for motion prediction in self-driving vehicles that learns environment representations via redundancy reduction. Our first type of redundancy reduction is induced by an interna…

Decodermotion predictionPredictionRepresentation Learning+3