paper-with-me

Papers

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility

2025-09-29 · Yutong Hao, Chen Chen, Ajmal Saeed Mian, Chang Xu, Daochang Liu arxiv

Diffusion models can generate realistic videos, but existing methods rely on implicitly learning physical reasoning from large-scale text-video datasets, which is costly, difficult to scale, and still prone to producing implausible motions that violate fundamental physical laws. We introduce a training-free framework that improves physical plausibility at inference time by explicitly reasoning about implausibility and guiding the generation away from it. Specifically, we employ a lightweight physics-aware reasoning pipeline to construct counterfactual prompts that deliberately encode physics-violating behaviors. Then, we propose a novel Synchronized Decoupled Guidance (SDG) strategy, which leverages these prompts through synchronized directional normalization to counteract lagged suppression and trajectory-decoupled denoising to mitigate cumulative trajectory bias, ensuring that implausible content is suppressed immediately and consistently throughout denoising. Experiments across different physical domains show that our approach substantially enhances physical fidelity while maintaining photorealism, despite requiring no additional training. Ablation studies confirm the complementary effectiveness of both the physics-aware reasoning component and SDG. In particular, the aforementioned two designs of SDG are also individually validated to contribute critically to the suppression of implausible content and the overall gains in physical plausibility. This establishes a new and plug-and-play physics-aware paradigm for video generation.

📄 PDF Abstract BibTeX arXiv:2509.24702

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

PhysMRV: Physical Memory Retrieval and Verification for Physics Plausibility Reasoning

2026-07-11 · Wenyuan Wang, Lianyu Hu, Hao Wang, Yang Liu arxiv

Video-language models (VLMs) have achieved remarkable performance on video understanding and visual question answering, yet they remain unreliable in reasoning about physical plausibility, where understanding object inte…

Physical Commonsense ReasoningVisual Question Answering

Morpheus: Benchmarking Physical Reasoning of Video Generative Models with Real Physical Experiments

2025-04-03 · Chenyu Zhang, Daniil Cherniavskii, Andrii Zadaianchuk, Antonios Tragoudaras 외

Recent advances in image and video generation raise hopes that these models possess world modeling capabilities, the ability to generate realistic, physically plausible videos. This could revolutionize applications in ro…

Physical Commonsense ReasoningVideo Generation

CausalMotion: Structured Physical Reasoning as Keyframe and Trajectory Guidance for Training-Free Video Generation

2026-06-12 · Sihan Zhuang, Xinyuan Chen, Tianfan Xue, Yaohui Wang arxiv

Recent advances in diffusion-based video generation have significantly improved visual quality and short-term temporal coherence. However, existing methods still struggle to produce videos with physically consistent and …

Video Generation

PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation

2025-09-24 · Chen Wang, Chuhao Chen, Yiming Huang, Zhiyang Dou 외 arxiv

Existing video generation models excel at producing photo-realistic videos from text or images, but often lack physical plausibility and 3D controllability. To overcome these limitations, we introduce PhysCtrl, a novel f…

Video Generation

PhyRPR: Training-Free Physics-Constrained Video Generation

2026-01-14 · Yibo Zhao, Hengjia Li, Xiaofei He, Boxi Wu arxiv

Recent diffusion-based video generation models can synthesize visually plausible videos, yet they often struggle to satisfy physical constraints. A key reason is that most existing approaches remain single-stage: they en…

Video Generation