paper-with-me

홈 › Papers

Anchoring and Rescaling Attention for Semantically Coherent Inbetweening

2026-03-18 · Tae Eun Choi, Sumin Shim, Junhyeok Kim, Seong Jae Hwang arxiv

Generative inbetweening (GI) seeks to synthesize realistic intermediate frames between the first and last keyframes beyond mere interpolation. As sequences become sparser and motions larger, previous GI models struggle with inconsistent frames with unstable pacing and semantic misalignment. Since GI involves fixed endpoints and numerous plausible paths, this task requires additional guidance gained from the keyframes and text to specify the intended path. Thus, we give semantic and temporal guidance from the keyframes and text onto each intermediate frame through Keyframe-anchored Attention Bias. We also better enforce frame consistency with Rescaled Temporal RoPE, which allows self-attention to attend to keyframes more faithfully. TGI-Bench, the first benchmark specifically designed for text-conditioned GI evaluation, enables challenge-targeted evaluation to analyze GI models. Without additional training, our method achieves state-of-the-art frame consistency, semantic fidelity, and pace stability for both short and long sequences across diverse challenges.

📄 PDF Abstract BibTeX arXiv:2603.17651

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion

2025-10-14 · Jungbin Cho, Minsu Kim, Jisoo Kim, Ce Zheng 외 arxiv

Human motion is inherently diverse and semantically rich, while also shaped by the surrounding scene. However, existing motion generation approaches fail to generate semantically diverse motion while simultaneously respe…

Motion Prior Distillation in Time Reversal Sampling for Generative Inbetweening

2026-02-13 · Wooseok Jeon, Seunghyun Shin, Dongmin Shin, Hae-Gon Jeon arxiv

Recent progress in image-to-video (I2V) diffusion models has significantly advanced the field of generative inbetweening, which aims to generate semantically plausible frames between two keyframes. In particular, inferen…

StructInbet: Integrating Explicit Structural Guidance into Inbetween Frame Generation

2025-07-15 · Zhenglin Pan, Haoran Xie arxiv

In this paper, we propose StructInbet, an inbetweening system designed to generate controllable transitions over explicit structural guidance. StructInbet introduces two key contributions. First, we propose explicit stru…

Generative Inbetweening through Frame-wise Conditions-Driven Video Generation

2024-12-16 · CVPR 2025 1 · Tianyi Zhu, Dongwei Ren, Qilong Wang, Xiaohe Wu 외

Generative inbetweening aims to generate intermediate frame sequences by utilizing two key frames as input. Although remarkable progress has been made in video generation models, generative inbetweening still faces chall…

Video Generation

Out of Distribution Detection via Neural Network Anchoring

2022-07-08 · Rushil Anirudh, Jayaraman J. Thiagarajan

Our goal in this paper is to exploit heteroscedastic temperature scaling as a calibration strategy for out of distribution (OOD) detection. Heteroscedasticity here refers to the fact that the optimal temperature paramete…

Out-of-Distribution DetectionOut of Distribution (OOD) Detection