paper-with-me

홈 › Papers

Learning to Film From Professional Human Motion Videos

2019-06-01 · CVPR 2019 6 · Chong Huang, Chuan-En Lin, Zhenyu Yang, Yan Kong, Peng Chen, Xin Yang, Kwang-Ting Cheng

We investigate the problem of 6 degrees of freedom (DOF) camera planning for filming professional human motion videos using a camera drone. Existing methods either plan motions for only a pan-tilt-zoom (PTZ) camera, or adopt ad-hoc solutions without carefully considering the impact of video contents and previous camera motions on the future camera motions. As a result, they can hardly achieve satisfactory results in our drone cinematography task. In this study, we propose a learning-based framework which incorporates the video contents and previous camera motions to predict the future camera motions that enable the capture of professional videos. Specifically, the inputs of our framework are video contents which are represented using subject-related feature based on 2D skeleton and scene-related features extracted from background RGB images, and camera motions which are represented using optical flows. The correlation between the inputs and output future camera motions are learned via a sequence-to-sequence convolutional long short-term memory (Seq2Seq ConvLSTM) network from a large set of video clips. We deploy our approach to a real drone cinematography system by first predicting the future camera motions, and then converting them to the drone's control commands via an odometer. Our experimental results on extensive datasets and showcases exhibit significant improvements in our approach over conventional baselines and our approach can successfully mimic the footage of a professional cameraman.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Building a Precise Video Language with Human-AI Oversight

2026-04-22 · Zhiqiu Lin, Chancharik Mitra, Siyuan Cen, Isaac Li 외 arxiv

Video-language models (VLMs) learn to reason about the dynamic visual world through natural language. We introduce a suite of open datasets, benchmarks, and recipes for scalable oversight that enable precise video captio…

Video CaptioningVideo GenerationText Generation

One-Shot Imitation Filming of Human Motion Videos

2019-12-23 · Chong Huang, Yuanjie Dang, Peng Chen, Xin Yang 외

Imitation learning has been applied to mimic the operation of a human cameraman in several autonomous cinematography systems. To imitate different filming styles, existing methods train multiple models, where each model …

Imitation LearningStyle Transfer

World Narrative Model for Highly Controllable Video Generation: A Paradigm Shift from Pixel Sampling to Physical World Orchestration

2026-06-30 · Ye Chen, Xuanhong Chen, Yupeng Zhu, Liming Tan 외 arxiv

The fundamental obstacle to industrial grade video generation is the lack of controllability: existing models treat video as a pixel distribution sampling problem, bypassing the explicit, instance level $4D$ $(3D + T)$ p…

Video Generation

Stable Cinemetrics : Structured Taxonomy and Evaluation for Professional Video Generation

2025-09-30 · Agneet Chatterjee, Rahim Entezari, Maksym Zhuravinskyi, Maksim Lapin 외 arxiv

Recent advances in video generation have enabled high-fidelity video synthesis from user provided prompts. However, existing models and benchmarks fail to capture the complexity and requirements of professional video gen…

Question GenerationVideo Generation

Aesthetics Driven Autonomous Time-Lapse Photography Generation by Virtual and Real Robots

2022-08-22 · Xiaobo Gao, Qi Kuang, Xin Jin, Bin Zhou 외

Time-lapse photography is employed in movies and promotional films because it can reflect the passage of time in a short time and strengthen the visual attraction. However, since it takes a long time and requires the sta…