paper-with-me

Papers

LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing

2026-06-25 · Xinyu Wang, Chongbo Zhao, Fangneng Zhan, Yue Ma arxiv

Streaming video editing has made rapid progress, yet practical deployment is still limited by two core issues: maintaining stable backgrounds and non-edited regions over time, and achieving the low latency required for real-time interactive scenarios. Meanwhile, recent streaming video generation methods are mostly developed for synthesis and cannot be directly applied to editing due to the strict preservation requirement and region-specific control. In this work, we present a novel streaming video editing framework that performs causal, frame-by-frame editing with strong content preservation and real-time responsiveness. Our key design is a three-stage distillation pipeline that progressively transfers editing capability from a powerful bidirectional foundation model to an efficient unidirectional streaming editor, enabling stable long-horizon edits without sacrificing visual fidelity. To further support real-time deployment, we introduce an AR-oriented mask cache that reuses region-related computation across frames, substantially reducing redundant processing and accelerating inference. Finally, we establish a dedicated benchmark for streaming video editing. Extensive evaluations demonstrate that our method achieves state-of-the-art visual quality among streaming baselines while drastically boosting inference speed to 12.66 FPS, making it suitable for interactive and augmented reality applications.

📄 PDF Abstract BibTeX arXiv:2606.26740

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

Streaming Video Diffusion: Online Video Editing with Diffusion Models

2024-05-30 · Feng Chen, Zhen Yang, Bohan Zhuang, Qi Wu

We present a novel task called online video editing, which is designed to edit \textbf{streaming} frames while maintaining temporal consistency. Unlike existing offline video editing assuming all frames are pre-establish…

Video Editing

REST: Diffusion-based Real-time End-to-end Streaming Talking Head Generation via ID-Context Caching and Asynchronous Streaming Distillation

2025-12-12 · Haotian Wang, Yuzhe Weng, Jun Du, Haoran Xu 외 arxiv

Diffusion models have significantly advanced the field of talking head generation (THG). However, slow inference speeds and prevalent non-autoregressive paradigms severely constrain the application of diffusion-based THG…

Talking Head Generation

InstantViR: Real-Time Video Inverse Problem Solver with Distilled Diffusion Prior

2025-11-18 · Weimin Bai, Suzhe Xu, Yiwei Ren, Jinhua Hao 외 arxiv

Video inverse problems are fundamental to streaming, telepresence, and AR/VR, where high perceptual quality must coexist with tight latency constraints. Diffusion-based priors currently deliver state-of-the-art reconstru…

Video ReconstructionVideo Restoration

FlashVSR: Towards Real-Time Diffusion-Based Streaming Video Super-Resolution

2025-10-14 · Junhao Zhuang, Shi Guo, Xin Cai, Xiaohui Li 외 arxiv

Diffusion models have recently advanced video restoration, but applying them to real-world video super-resolution (VSR) remains challenging due to high latency, prohibitive computation, and poor generalization to ultra-h…

Video Super-ResolutionVideo Restoration

Semantic-Aware Adaptive Video Streaming Using Latent Diffusion Models for Wireless Networks

2025-02-08 · Zijiang Yan, Jianhua Pei, Hongda Wu, Hina Tabassum 외

This paper proposes a novel framework for real-time adaptive-bitrate video streaming by integrating latent diffusion models (LDMs) within the FFmpeg techniques. This solution addresses the challenges of high bandwidth us…

DenoisingVideo Frame InterpolationVideo Reconstruction