paper-with-me

홈 › Papers

PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolution

2026-05-25 · Wenxue Li, Jingjing Ren, Peng Zhang, Tian Ye, Daiguo Zhou, Jian Luan, Lei Zhu arxiv

High-resolution video generation faces a coupled bottleneck of optimization instability and prohibitive computational costs. The massive expansion of the token sequence not only biases optimization toward local textures at the expense of global coherence, leading to structural collapse, but also imposes prohibitive training costs and severe inference latency. To address this, we propose PixelWizard, a framework that hierarchically decouples global structure modeling from fine-grained detail synthesis. PixelWizard first establishes a compact spatiotemporal anchor to concentrate dense structural priors, which then guides fine-grained generation at high resolution. This mitigates the local optimization bias to ensure structural stability without compromising high-frequency details. Leveraging this structural stability, we introduce Noise-Span Aligned Shortcut Training to break the inference bottleneck. By explicitly modeling the step size, this mechanism allows the model to traverse the generation trajectory with large steps. Crucially, we incorporate Exponential Index-Biased Sampling and Adaptive Noise-Span Calibration to align optimization with the shifted noise schedules of high-resolution grids, ensuring robust few-step inference without incurring the heavy overhead of distillation. Extensive experiments demonstrate that PixelWizard achieves superior visual quality while accelerating the generative sampling of native 2K/4K videos by over 10x.

📄 PDF Abstract BibTeX arXiv:2605.25801

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References

2026-03-26 · Jiahao Wang, Hualian Sheng, Sijia Cai, Yuxiao Yang 외 arxiv

Identity-preserving video generation offers powerful tools for creative expression, allowing users to customize videos featuring their beloved characters. However, prevailing methods are typically designed and optimized …

Reinforcement LearningVideo Generation

HALLELUAI: A Hallucination-Aware AI System for Ultra-Realistic Image-to-Video Generation at Scale

2026-07-25 · Aniket Sakpal, Yang Jiang, Rouzbeh Davoudi, Shayan Hassantabar 외 arxiv

AI-generated video is increasingly used across marketing, product storytelling, and creative workflows, yet automated; high-precision quality control remains a major constraint to scaling production. We present HALLELUAI…

Video Generation

UltraGen: High-Resolution Video Generation with Hierarchical Attention

2025-10-21 · Teng Hu, Jiangning Zhang, Zihan Su, Ran Yi arxiv

Recent advances in video generation have made it possible to produce visually compelling videos, with wide-ranging applications in content creation, entertainment, and virtual reality. However, most existing diffusion tr…

Video Generation

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts

2026-02-12 · Chen Zhao, Jiawei Chen, Hongyu Li, Zhuoliang Kang 외 arxiv

Recent advances in video diffusion models have significantly improved visual quality, yet ultra-high-resolution (UHR) video generation remains a formidable challenge due to the compounded difficulties of motion modeling,…

Video Generation

Resolution-Agnostic Neural Compression for High-Fidelity Portrait Video Conferencing via Implicit Radiance Fields

2024-02-26 · Yifei Li, Xiaohong Liu, Yicong Peng, Guangtao Zhai 외

Video conferencing has caught much more attention recently. High fidelity and low bandwidth are two major objectives of video compression for video conferencing applications. Most pioneering methods rely on classic video…

Video Compression