paper-with-me

Papers

V-Trans4Style: Visual Transition Recommendation for Video Production Style Adaptation

2025-01-14 · Pooja Guhan, Tsung-Wei Huang, Guan-Ming Su, Subhadra Gopalakrishnan, Dinesh Manocha

We introduce V-Trans4Style, an innovative algorithm tailored for dynamic video content editing needs. It is designed to adapt videos to different production styles like documentaries, dramas, feature films, or a specific YouTube channel's video-making technique. Our algorithm recommends optimal visual transitions to help achieve this flexibility using a more bottom-up approach. We first employ a transformer-based encoder-decoder network to learn recommending temporally consistent and visually seamless sequences of visual transitions using only the input videos. We then introduce a style conditioning module that leverages this model to iteratively adjust the visual transitions obtained from the decoder through activation maximization. We demonstrate the efficacy of our method through experiments conducted on our newly introduced AutoTransition++ dataset. It is a 6k video version of AutoTransition Dataset that additionally categorizes its videos into different production style categories. Our encoder-decoder model outperforms the state-of-the-art transition recommendation method, achieving improvements of 10% to 80% in Recall@K and mean rank values over baseline. Our style conditioning module results in visual transitions that improve the capture of the desired video production style characteristics by an average of around 12% in comparison to other methods when measured with similarity metrics. We hope that our work serves as a foundation for exploring and understanding video production styles further.

📄 PDF Abstract BibTeX arXiv:2501.07983

Code (0)

등록된 구현이 없습니다.

Tasks

Decoder

Similar Papers 제목 키워드 기반

SOYO: A Tuning-Free Approach for Video Style Morphing via Style-Adaptive Interpolation in Diffusion Models

2025-03-10 · Haoyu Zheng, Qifan Yu, Binghe Yu, Yang Dai 외

Diffusion models have achieved remarkable progress in image and video stylization. However, most existing methods focus on single-style transfer, while video stylization involving multiple styles necessitates seamless tr…

Style Transfer

Edit3K: Universal Representation Learning for Video Editing Components

2024-03-24 · Xin Gu, Libo Zhang, Fan Chen, Longyin Wen 외

This paper focuses on understanding the predominant video creation pipeline, i.e., compositional video editing with six main types of editing components, including video effects, animation, transition, filter, sticker, a…

Representation LearningRetrievalVideo Editing

Real-time Localized Photorealistic Video Style Transfer

2020-10-20 · Xide Xia, Tianfan Xue, Wei-Sheng Lai, Zheng Sun 외

We present a novel algorithm for transferring artistic styles of semantically meaningful local regions of an image onto local regions of a target video while preserving its photorealism. Local regions may be selected eit…

Style TransferVideo SegmentationVideo Semantic SegmentationVideo Style Transfer

AutoTransition: Learning to Recommend Video Transition Effects

2022-07-27 · Yaojie Shen, Libo Zhang, Kai Xu, Xiaojie Jin

Video transition effects are widely used in video editing to connect shots for creating cohesive and visually appealing videos. However, it is challenging for non-professionals to choose best transitions due to the lack …

RetrievalVideo Editing

Generative Disco: Text-to-Video Generation for Music Visualization

2023-04-17 · Vivian Liu, Tao Long, Nathan Raw, Lydia Chilton

Visuals can enhance our experience of music, owing to the way they can amplify the emotions and messages conveyed within it. However, creating music visualization is a complex, time-consuming, and resource-intensive proc…

Text-to-Video GenerationVideo Generation