paper-with-me

홈 › Papers

Pre-training for Action Recognition with Automatically Generated Fractal Datasets

2024-11-26 · Davyd Svyezhentsev, George Retsinas, Petros Maragos

In recent years, interest in synthetic data has grown, particularly in the context of pre-training the image modality to support a range of computer vision tasks, including object classification, medical imaging etc. Previous work has demonstrated that synthetic samples, automatically produced by various generative processes, can replace real counterparts and yield strong visual representations. This approach resolves issues associated with real data such as collection and labeling costs, copyright and privacy. We extend this trend to the video domain applying it to the task of action recognition. Employing fractal geometry, we present methods to automatically produce large-scale datasets of short synthetic video clips, which can be utilized for pre-training neural models. The generated video clips are characterized by notable variety, stemmed by the innate ability of fractals to generate complex multi-scale structures. To narrow the domain gap, we further identify key properties of real videos and carefully emulate them during pre-training. Through thorough ablations, we determine the attributes that strengthen downstream results and offer general guidelines for pre-training with synthetic videos. The proposed approach is evaluated by fine-tuning pre-trained models on established action recognition datasets HMDB51 and UCF101 as well as four other video benchmarks related to group action recognition, fine-grained action recognition and dynamic scenes. Compared to standard Kinetics pre-training, our reported results come close and are even superior on a portion of downstream datasets. Code and samples of synthetic videos are available at https://github.com/davidsvy/fractal_video .

📄 PDF Abstract BibTeX arXiv:2411.17584

Code (1)

davidsvy/fractal_video 공식 구현 pytorch

Tasks

Action RecognitionFine-grained Action Recognition

Similar Papers 제목 키워드 기반

How to Sample High Quality 3D Fractals for Action Recognition Pre-Training?

2026-02-12 · Marko Putak, Thomas B. Moeslund, Joakim Bruslund Haurum arxiv

Synthetic datasets are being recognized in the deep learning realm as a valuable alternative to exhaustively labeled real data. One such synthetic data generation method is Formula Driven Supervised Learning (FDSL), whic…

Synthetic Data GenerationAction Recognition

Pre-training without Natural Images

2021-01-21 · Hirokatsu Kataoka, Kazushige Okayasu, Asato Matsumoto, Eisuke Yamagata 외

Is it possible to use convolutional neural networks pre-trained without any natural images to assist natural image understanding? The paper proposes a novel concept, Formula-driven Supervised Learning. We automatically g…

Preparation of Fractal-Inspired Computational Architectures for Advanced Large Language Model Analysis

2025-11-10 · Yash Mittal, Dmitry Ignatov, Radu Timofte arxiv

This paper proposes FractalNet, a framework based on fractal design principles that automatically generates and evaluates convolutional neural network (CNN) architectures using recursive template patterns. Rather than re…

Neural Architecture SearchImage Classification

Multiscale Fractal Analysis on EEG Signals for Music-Induced Emotion Recognition

2020-10-30 · Kleanthis Avramidis, Athanasia Zlatintsi, Christos Garoufis, Petros Maragos

Emotion Recognition from EEG signals has long been researched as it can assist numerous medical and rehabilitative applications. However, their complex and noisy structure has proven to be a serious barrier for tradition…

EEGElectroencephalogram (EEG)Emotion ClassificationEmotion Recognition

Learning Fractals by Gradient Descent

2023-03-14 · Cheng-Hao Tu, Hong-You Chen, David Carlyn, Wei-Lun Chao

Fractals are geometric shapes that can display complex and self-similar patterns found in nature (e.g., clouds and plants). Recent works in visual recognition have leveraged this property to create random fractal images …