paper-with-me

Papers Video Editing

“Video Editing” 태그가 달린 논문 346편 · 필터 해제

DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing

2025-06-26 · Lingling Cai, Kang Zhao, Hangjie Yuan, Xiang Wang 외

The advent of Video Diffusion Transformers (Video DiTs) marks a milestone in video generation. However, directly applying existing video editing methods to Video DiTs often incurs substantial computational overhead, due …

Video EditingVideo Generation

Let Your Video Listen to Your Music!

2025-06-23 · Xinyu Zhang, Dong Gong, Zicheng Duan, Anton Van Den Hengel 외

Aligning the rhythm of visual motion in a video with a given music track is a practical need in multimedia production, yet remains an underexplored task in autonomous video editing. Effective alignment between motion and…

GPUMusic GenerationRhythmVideo Editing+1

Causally Steered Diffusion for Automated Video Counterfactual Generation

2025-06-17 · Nikos Spyrou, Athanasios Vlontzos, Paraskevas Pegios, Thomas Melistas 외

Adapting text-to-image (T2I) latent diffusion models for video editing has shown strong visual fidelity and controllability, but challenges remain in maintaining causal relationships in video content. Edits affecting cau…

counterfactualVideo EditingVideo Generation

LoRA-Edit: Controllable First-Frame-Guided Video Editing via Mask-Aware LoRA Fine-Tuning

2025-06-11 · Chenjian Gao, Lihe Ding, Xin Cai, Zhanpeng Huang 외

Video editing using diffusion models has achieved remarkable results in generating high-quality edits for videos. However, current methods often rely on large-scale pretraining, limiting flexibility for specific edits. F…

Video Editing

RoboSwap: A GAN-driven Video Diffusion Framework For Unsupervised Robot Arm Swapping

2025-06-10 · Yang Bai, Liudi Yang, George Eskandar, Fengyi Shen 외

Recent advancements in generative models have revolutionized video synthesis and editing. However, the scarcity of diverse, high-quality datasets continues to hinder video-conditioned robotic learning, limiting cross-pla…

Video Editing

Super Encoding Network: Recursive Association of Multi-Modal Encoders for Video Understanding

2025-06-09 · BoYu Chen, Siran Chen, Kunchang Li, Qinglin Xu 외

Video understanding has been considered as one critical step towards world modeling, which is an important long-term problem in AI research. Recently, multi-modal foundation models have shown such potential via large-sca…

Contrastive LearningVideo EditingVideo Understanding

FADE: Frequency-Aware Diffusion Model Factorization for Video Editing

2025-06-06 · CVPR 2025 1 · Yixuan Zhu, Haolin Wang, Shilin Ma, Wenliang Zhao 외

Recent advancements in diffusion frameworks have significantly enhanced video editing, achieving high fidelity and strong alignment with textual prompts. However, conventional approaches using image diffusion models fall…

Video Editing

FlowDirector: Training-Free Flow Steering for Precise Text-to-Video Editing

2025-06-05 · Guangzhao Li, Yanming Yang, Chenxi Song, Chi Zhang

Text-driven video editing aims to modify video content according to natural language instructions. While recent training-free approaches have made progress by leveraging pre-trained diffusion models, they typically rely …

Text-to-Video EditingVideo Editing

FullDiT2: Efficient In-Context Conditioning for Video Diffusion Transformers

2025-06-04 · Xuanhua He, Quande Liu, Zixuan Ye, Weicai Ye 외

Fine-grained and efficient controllability on video diffusion transformers has raised increasing desires for the applicability. Recently, In-context Conditioning emerged as a powerful paradigm for unified conditional vid…

Video EditingVideo Generation

MiniMax-Remover: Taming Bad Noise Helps Video Object Removal

2025-05-30 · Bojia Zi, Weixuan Peng, Xianbiao Qi, Jianan Wang 외

Recent advances in video diffusion models have driven rapid progress in video editing techniques. However, video object removal, a critical subtask of video editing, remains challenging due to issues such as hallucinated…

Video EditingVideo Generation

Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing

2025-05-29 · Tongtong Su, Chengyu Wang, Jun Huang, Dongming Lu

Appearance editing according to user needs is a pivotal task in video editing. Existing text-guided methods often lead to ambiguities regarding user intentions and restrict fine-grained control over editing specific aspe…

Optical Flow EstimationVideo EditingVideo Restoration

Video Editing for Audio-Visual Dubbing

2025-05-29 · Binyamin Manela, Sharon Gannot, Ethan Fetyaya

Visual dubbing, the synchronization of facial movements with new speech, is crucial for making content accessible across different languages, enabling broader global reach. However, current methods face significant limit…

Video Editing

TDVE-Assessor: Benchmarking and Evaluating the Quality of Text-Driven Video Editing with LMMs

2025-05-26 · Juntong Wang, Jiarui Wang, Huiyu Duan, Guangtao Zhai 외

Text-driven video editing is rapidly advancing, yet its rigorous evaluation remains challenging due to the absence of dedicated video quality assessment (VQA) models capable of discerning the nuances of editing quality. …

BenchmarkingLarge Language ModelVideo EditingVideo Quality Assessment+1

SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation

2025-05-25 · Shenggan Cheng, Yuanxin Wei, Lansong Diao, Yong liu 외

Leveraging the diffusion transformer (DiT) architecture, models like Sora, CogVideoX and Wan have achieved remarkable progress in text-to-video, image-to-video, and video editing tasks. Despite these advances, diffusion-…

Video EditingVideo Generation

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing

2025-05-24 · Weihan Xu, Yimeng Ma, Jingyue Huang, Yang Li 외

Short videos are an effective tool for promoting contents and improving knowledge accessibility. While existing extractive video summarization methods struggle to produce a coherent narrative, existing abstractive method…

Language ModelingLanguage ModellingLarge Language ModelRetrieval+2

From Shots to Stories: LLM-Assisted Video Editing with Unified Language Representations

2025-05-18 · Yuzhi Li, Haojun Xu, Fang Tian

Large Language Models (LLMs) and Vision-Language Models (VLMs) have demonstrated remarkable reasoning and generalization capabilities in video understanding; however, their application in video editing remains largely un…

Video EditingVideo Understanding

DAPE: Dual-Stage Parameter-Efficient Fine-Tuning for Consistent Video Editing with Diffusion Models

2025-05-11 · Junhao Xia, Chaoyang Zhang, Yecheng Zhang, Chengyang Zhou 외

Video generation based on diffusion models presents a challenging multimodal task, with video editing emerging as a pivotal direction in this field. Recent video editing approaches primarily fall into two categories: tra…

parameter-efficient fine-tuningVideo AlignmentVideo EditingVideo Generation

Video Forgery Detection for Surveillance Cameras: A Review

2025-05-04 · Noor B. Tayfor, Tarik A. Rashid, Shko M. Qader, Bryar A. Hassan 외

The widespread availability of video recording through smartphones and digital devices has made video-based evidence more accessible than ever. Surveillance footage plays a crucial role in security, law enforcement, and …

Frame Duplication DetectionMisinformationVideo Editing

A Rusty Link in the AI Supply Chain: Detecting Evil Configurations in Model Repositories

2025-05-02 · Ziqi Ding, QiAn Fu, Junchen Ding, Gelei Deng 외

Recent advancements in large language models (LLMs) have spurred the development of diverse AI applications from code generation and video editing to text generation; however, AI supply chains such as Hugging Face, which…

Code GenerationText GenerationVideo Editing

Controllable Weather Synthesis and Removal with Video Diffusion Models

2025-05-01 · Chih-Hao Lin, Zian Wang, Ruofan Liang, Yuxuan Zhang 외

Generating realistic and controllable weather effects in videos is valuable for many applications. Physics-based weather simulation requires precise reconstructions that are hard to scale to in-the-wild videos, while cur…

Video Editing
1–20 / 346 다음 →