paper-with-me

Papers

CineLOG: A Training Free Approach for Cinematic Long Video Generation

2025-12-13 · Zahra Dehghanian, Morteza Abolghasemi, Hamid Beigy, Hamid R. Rabiee arxiv

Controllable video synthesis is a central challenge in computer vision, yet current models struggle with fine grained control beyond textual prompts, particularly for cinematic attributes like camera trajectory and genre. Existing datasets often suffer from severe data imbalance, noisy labels, or a significant simulation to real gap. To address this, we introduce CineLOG, a new dataset of 5,000 high quality, balanced, and uncut video clips. Each entry is annotated with a detailed scene description, explicit camera instructions based on a standard cinematic taxonomy, and genre label, ensuring balanced coverage across 17 diverse camera movements and 15 film genres. We also present our novel pipeline designed to create this dataset, which decouples the complex text to video (T2V) generation task into four easier stages with more mature technology. To enable coherent, multi shot sequences, we introduce a novel Trajectory Guided Transition Module that generates smooth spatio-temporal interpolation. Extensive human evaluations show that our pipeline significantly outperforms SOTA end to end T2V models in adhering to specific camera and screenplay instructions, while maintaining professional visual quality. All codes and data are available at https://cine-log.pages.dev.

📄 PDF Abstract BibTeX arXiv:2512.12209

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

CineWeaver: Training-Free Reference-Controllable Multi-Shot Long Video Generation for Cinematic Storytelling

2026-07-29 · Yuyang Huang, Yabo Chen, Wenrui Dai, Ziyang Zheng 외 arxiv

Cinematic video generation is challenging for text-to-video diffusion models due to concurrent requirements on multi-shot generation, fine-grained controllability over characters and scenes, and long-form generation acro…

Video Generation

CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video Generation

2026-02-06 · Kaiyi Huang, Yukun Huang, Yu Li, Jianhong Bai 외 arxiv

Cinematic video production requires control over scene-subject composition and camera movement, but live-action shooting remains costly due to the need for constructing physical sets. To address this, we introduce the ta…

Text-to-Video Generation

Automatic Funny Scene Extraction from Long-form Cinematic Videos

2026-02-17 · Sibendu Paul, Haotian Jiang, Caren Chen arxiv

Automatically extracting engaging and high-quality humorous scenes from cinematic titles is pivotal for creating captivating video previews and snackable content, boosting user engagement on streaming platforms. Long-for…

Scene Segmentation

A Benchmark and Multi-Agent System for Instruction-driven Cinematic Video Compilation

2026-04-12 · Peixuan Zhang, Chang Zhou, Ziyuan Zhang, Hualuo Liu 외 arxiv

The surging demand for adapting long-form cinematic content into short videos has motivated the need for versatile automatic video compilation systems. However, existing compilation methods are limited to predefined task…

ReCA: Multi-Shot Long Video Extrapolation via Recursive Context Allocation

2026-05-26 · Akide Liu, Jinbo Xing, Chaojie Mao, Ye Li 외 arxiv

Minute-scale cinematic video generation is a central challenge for generative video models. Existing paradigms address only fragments of this challenge: single-shot extrapolation preserves an anchor but lacks cinematic s…

Video Generation