Condensing Action Segmentation Datasets via Generative Network Inversion
This work presents the first condensation approach for procedural video datasets used in temporal action segmentation. We propose a condensation framework that leverages generative prior learned from the dataset and network inversion to condense data into compact latent codes with significant storage reduced across temporal and channel aspects. Orthogonally, we propose sampling diverse and representative action sequences to minimize video-wise redundancy. Our evaluation on standard benchmarks demonstrates consistent effectiveness in condensing TAS datasets and achieving competitive performances. Specifically, on the Breakfast dataset, our approach reduces storage by over 500$\times$ while retaining 83% of the performance compared to training with the full dataset. Furthermore, when applied to a downstream incremental learning task, it yields superior performance compared to the state-of-the-art.
Code (0)
등록된 구현이 없습니다.
Tasks
Action SegmentationIncremental LearningTemporal Action SegmentationSimilar Papers 제목 키워드 기반
Zero-shot Segmentation of Skin Conditions: Erythema with Edit-Friendly Inversion
This study proposes a zero-shot image segmentation framework for detecting erythema (redness of the skin) using edit-friendly inversion in diffusion models. The method synthesizes reference images of the same patient tha…
Image SegmentationRGI: robust GAN-inversion for mask-free image inpainting and unsupervised pixel-wise anomaly detection
Generative adversarial networks (GANs), trained on a large-scale image dataset, can be a good approximator of the natural image manifold. GAN-inversion, using a pre-trained generator as a deep generative prior, is a prom…
Anomaly DetectionImage InpaintingImage RestorationWeighted Distance Nearest Neighbor Condensing
The problem of nearest neighbor condensing has enjoyed a long history of study, both in its theoretical and practical aspects. In this paper, we introduce the problem of weighted distance nearest neighbor condensing, whe…
Generalization BoundsSNR-Edit: Structure-Aware Noise Rectification for Inversion-Free Flow-Based Editing
Inversion-free image editing using flow-based generative models challenges the prevailing inversion-based pipelines. However, existing approaches rely on fixed Gaussian noise to construct the source trajectory, leading t…
Image EditingPrivacy Vulnerability of Split Computing to Data-Free Model Inversion Attacks
Mobile edge devices see increased demands in deep neural networks (DNNs) inference while suffering from stringent constraints in computing resources. Split computing (SC) emerges as a popular approach to the issue by exe…