paper-with-me

홈 › Papers

Temporally Coherent Interpretations for Long Videos Using Pattern Theory

2015-06-01 · CVPR 2015 6 · Fillipe Souza, Sudeep Sarkar, Anuj Srivastava, Jingyong Su

Graph-theoretical methods have successfully provided semantic and structural interpretations of images and videos. A recent paper introduced a pattern-theoretic approach that allows construction of flexible graphs for representing interactions of actors with objects and inference is accomplished by an efficient annealing algorithm. Actions and objects are termed generators and their interactions are termed bonds; together they form high-probability configurations, or interpretations, of observed scenes. This work and other structural methods have generally been limited to analyzing short videos involving isolated actions. Here we provide an extension that uses additional temporal bonds across individual actions to enable semantic interpretations of longer videos. Longer temporal connections improve scene interpretations as they help discard (temporally) local solutions in favor of globally superior ones. Using this extension, we demonstrate improvements in understanding longer videos, compared to individual interpretations of non-overlapping time segments. We verified the success of our approach by generating interpretations for more than 700 video segments from the YouCook data set, with intricate videos that exhibit cluttered background, scenarios of occlusion, viewpoint variations and changing conditions of illumination. Interpretations for long video segments were able to yield performance increases of about 70 and, in addition, proved to be more robust to different severe scenarios of classification errors.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Temporally coherent completion of dynamic video

2016-11-01 · J.-B. Huang, S. B. Kang, N. Ahuja, J. Kopf

We present an automatic video completion algorithm that synthesizes missing regions in videos in a temporally coherent fashion. Our algorithm can handle dynamic scenes captured using a moving camera. State-of-the-art app…

Optical Flow EstimationVideo Inpainting

MonoClothCap: Towards Temporally Coherent Clothing Capture from Monocular RGB Video

2020-09-22 · Donglai Xiang, Fabian Prada, Chenglei Wu, Jessica Hodgins

We present a method to capture temporally coherent dynamic clothing deformation from a monocular RGB video input. In contrast to the existing literature, our method does not require a pre-scanned personalized mesh templa…

Surface Reconstructionvalid

Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations

2025-11-25 · Chao Wang, Chengan Che, Xinyue Chen, Sophia Tsoka 외 arxiv

Counterfactual explanations (CFEs) are minimal and semantically meaningful modifications of the input of a model that alter the model predictions. They highlight the decisive features the model relies on, providing contr…

Emotion ClassificationAction Classification

FlowZero: Zero-Shot Text-to-Video Synthesis with LLM-Driven Dynamic Scene Syntax

2023-11-27 · Yu Lu, Linchao Zhu, Hehe Fan, Yi Yang

Text-to-video (T2V) generation is a rapidly growing research area that aims to translate the scenes, objects, and actions within complex video text into a sequence of coherent visual frames. We present FlowZero, a novel …

Video Generation

A Generative Framework for Probabilistic, Spatiotemporally Coherent Downscaling of Climate Simulation

2024-12-19 · Jonathan Schmidt, Luca Schmidt, Felix Strnad, Nicole Ludwig 외

Local climate information is crucial for impact assessment and decision-making, yet coarse global climate simulations cannot capture small-scale phenomena. Current statistical downscaling methods infer these phenomena as…

Decision Making