paper-with-me

홈 › Papers

MambaOVSR: Multiscale Fusion with Global Motion Modeling for Chinese Opera Video Super-Resolution

2025-11-09 · Hua Chang, Xin Xu, Wei Liu, Wei Wang, Xin Yuan, Kui Jiang arxiv

Chinese opera is celebrated for preserving classical art. However, early filming equipment limitations have degraded videos of last-century performances by renowned artists (e.g., low frame rates and resolution), hindering archival efforts. Although space-time video super-resolution (STVSR) has advanced significantly, applying it directly to opera videos remains challenging. The scarcity of datasets impedes the recovery of high frequency details, and existing STVSR methods lack global modeling capabilities, compromising visual quality when handling opera's characteristic large motions. To address these challenges, we pioneer a large scale Chinese Opera Video Clip (COVC) dataset and propose the Mamba-based multiscale fusion network for space-time Opera Video Super-Resolution (MambaOVSR). Specifically, MambaOVSR involves three novel components: the Global Fusion Module (GFM) for motion modeling through a multiscale alternating scanning mechanism, and the Multiscale Synergistic Mamba Module (MSMM) for alignment across different sequence lengths. Additionally, our MambaVR block resolves feature artifacts and positional information loss during alignment. Experimental results on the COVC dataset show that MambaOVSR significantly outperforms the SOTA STVSR method by an average of 1.86 dB in terms of PSNR. Dataset and Code will be publicly released.

📄 PDF Abstract BibTeX arXiv:2511.06172

Code (0)

등록된 구현이 없습니다.

Tasks

Space-time Video Super-resolution

Similar Papers 제목 키워드 기반

Towards Expressive Video Dubbing with Multiscale Multimodal Context Interaction

2024-12-25 · Yuan Zhao, Rui Liu, Gaoxiang Cong

Automatic Video Dubbing (AVD) generates speech aligned with lip motion and facial emotion from scripts. Recent research focuses on modeling multimodal context to enhance prosody expressiveness but overlooks two key issue…

Graph AttentionSentence

MSF-Mamba: Motion-aware State Fusion Mamba for Efficient Micro-Gesture Recognition

2025-10-12 · Deng Li, Jun Shao, Bohao Xing, Rong Gao 외 arxiv

Micro-gesture recognition (MGR) targets the identification of subtle and fine-grained human motions and requires accurate modeling of both long-range and local spatiotemporal dependencies. While CNNs are effective at cap…

Micro-gesture Recognition

Particle-based Multiscale Modeling of Calcium Puff Dynamics

2016-04-13

Intracellular calcium is regulated in part by the release of Ca$^{2+}$ ions from the endoplasmic reticulum via inositol-4,5-triphosphate receptor (IP$_3$R) channels (among other possibilities such as RyR and L-type calci…

MSTF: Multiscale Transformer for Incomplete Trajectory Prediction

2024-07-08 · Zhanwen Liu, Chao Li, Nan Yang, Yang Wang 외

Motion forecasting plays a pivotal role in autonomous driving systems, enabling vehicles to execute collision warnings and rational local-path planning based on predictions of the surrounding vehicles. However, prevalent…

Autonomous DrivingMissing ValuesMotion ForecastingPrediction+1

GFocal: A Global-Focal Neural Operator for Solving PDEs on Arbitrary Geometries

2025-08-06 · Fangzhi Fei, Jiaxin Hu, Qiaofeng Li, Zhenyu Liu arxiv

Transformer-based neural operators have emerged as promising surrogate solvers for partial differential equations, by leveraging the effectiveness of Transformers for capturing long-range dependencies and global correlat…