paper-with-me

홈 › Papers

TeleBoost: A Systematic Alignment Framework for High-Fidelity, Controllable, and Robust Video Generation

2026-02-07 · Yuanzhi Liang, Xuan'er Wu, Yirui Liu, Yijie Fang, Yizhen Fan, Ke Hao, Rui Li, Ruiying Liu, Ziqi Ni, Peng Yu, Yanbo Wang, Haibin Huang, Qizhen Weng, Chi Zhang, Xuelong Li arxiv

Post-training is the decisive step for converting a pretrained video generator into a production-oriented model that is instruction-following, controllable, and robust over long temporal horizons. This report presents a systematical post-training framework that organizes supervised policy shaping, reward-driven reinforcement learning, and preference-based refinement into a single stability-constrained optimization stack. The framework is designed around practical video-generation constraints, including high rollout cost, temporally compounding failure modes, and feedback that is heterogeneous, uncertain, and often weakly discriminative. By treating optimization as a staged, diagnostic-driven process rather than a collection of isolated tricks, the report summarizes a cohesive recipe for improving perceptual fidelity, temporal coherence, and prompt adherence while preserving the controllability established at initialization. The resulting framework provides a clear blueprint for building scalable post-training pipelines that remain stable, extensible, and effective in real-world deployment settings.

📄 PDF Abstract BibTeX arXiv:2602.07595

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningVideo Generation

Similar Papers 제목 키워드 기반

Improving and Assessing the Fidelity of Large Language Models Alignment to Online Communities

2024-08-18 · Minh Duc Chu, Zihao He, Rebecca Dorn, Kristina Lerman

Large language models (LLMs) have shown promise in representing individuals and communities, offering new ways to study complex social dynamics. However, effectively aligning LLMs with specific human groups and systemati…

Models Alignment

High Fidelity Text to Image Generation with Contrastive Alignment and Structural Guidance

2025-08-14 · Danyi Gao arxiv

This paper addresses the performance bottlenecks of existing text-driven image generation methods in terms of semantic alignment accuracy and structural consistency. A high-fidelity image generation method is proposed by…

Contrastive LearningImage Generation

DiTSinger: Scaling Singing Voice Synthesis with Diffusion Transformer and Implicit Alignment

2025-10-10 · Zongcai Du, Guilin Deng, Xiaofeng Guo, Xin Gao 외 arxiv

Recent progress in diffusion-based Singing Voice Synthesis (SVS) demonstrates strong expressiveness but remains limited by data scarcity and model scalability. We introduce a two-stage pipeline: a compact seed set of hum…

Cross-cultural value alignment frameworks for responsible AI governance: Evidence from China-West comparative analysis

2025-11-21 · Haijiang Liu, Jinguang Gu, Xun Wu, Daniel Hershcovich 외 arxiv

As Large Language Models (LLMs) increasingly influence high-stakes decision-making across global contexts, ensuring their alignment with diverse cultural values has become a critical governance challenge. This study pres…

Reinforcement Learning

An end-to-end agentic pipeline for smart contract translation and quality evaluation

2026-02-14 · Abhinav Goel, Chaitya Shah, Agostino Capponi, Alfio Gliozzo arxiv

We present an end-to-end framework for systematic evaluation of LLM-generated smart contracts from natural-language specifications. The system parses contractual text into structured schemas, generates Solidity code, and…