paper-with-me

Papers

Bridging Creative Intent and Visual Quality: Creator-Driven Recurrent Video Generation with Agentic Feedback Loops

2026-06-17 · Denis Savytski, Aiden Lei, Heding Liu, Warren Yang, Sihan Liang, Alexander Liu, Zhe Zhao arxiv

Generative AI has made content creation increasingly accessible, but many AI-generated videos lack narrative coherence and creative direction, issues that become more substantial at longer durations. Unlike coding, where AI generation benefits from reliable feedback and techniques such as recurrent self-improvement, video generation requires subjective feedback about plot, scenes, and narrative, which naturally motivates approaches that incorporate human creative direction. We introduce CHIEF, a human-AI co-creation video generation framework that places the creator at the center of human-in-the-loop iterative video refinement, and supports them by providing automatic subjective feedback. The creator incorporates their creative direction by driving each iteration, while their revisions are incorporated by a specialized refiner agent. The feedback loop is generated by persona-conditioned multimodal LLMs that watch generated videos and produce subjective critique from the audience perspectives, providing feedback that self-evaluation alone cannot capture. To test the effectiveness of our proposed framework, we work with high school and college students with no prior filmmaking experience to create videos, from short 1-minute videos to a complete short 10-minute film with a complicated plot.

📄 PDF Abstract BibTeX arXiv:2606.18591

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues

2025-12-02 · Zichen Liu, Yue Yu, Hao Ouyang, Qiuyu Wang 외 arxiv

We propose MagicQuill V2, a novel system that introduces a \textbf{layered composition} paradigm to generative image editing, bridging the gap between the semantic power of diffusion models and the granular control of tr…

Image Editing

Vidmento: Creating Video Stories Through Context-Aware Expansion With Generative Video

2026-01-29 · Catherine Yeh, Anh Truong, Mira Dontcheva, Bryan Wang arxiv

Video storytelling is often constrained by available material, limiting creative expression and leaving undesired narrative gaps. Generative video offers a new way to address these limitations by augmenting captured medi…

Can AI Be as Creative as Humans?

2024-01-03 · Haonan Wang, James Zou, Michael Mozer, Anirudh Goyal 외

Creativity serves as a cornerstone for societal progress and innovation. With the rise of advanced generative AI models capable of tasks once reserved for human creativity, the study of AI's creative potential becomes im…

VisionCreator: A Native Visual-Generation Agentic Model with Understanding, Thinking, Planning and Creation

2026-03-03 · Jinxiang Lai, Zexin Lu, Jiajun He, Rongwei Quan 외 arxiv

Visual content creation tasks demand a nuanced understanding of design conventions and creative workflows-capabilities challenging for general models, while workflow-based agents lack specialized knowledge for autonomous…

Reinforcement Learning

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

2025-05-21 · Jiaying Wu, Fanxiao Li, Min-Yen Kan, Bryan Hooi

The real-world impact of misinformation stems from the underlying misleading narratives that creators seek to convey. As such, interpreting misleading creator intent is essential for multimodal misinformation detection (…

ArticlesIntent DetectionMisinformation