paper-with-me

Papers

Learning Spatial-Temporal Implicit Neural Representations for Event-Guided Video Super-Resolution

2023-03-24 · CVPR 2023 1 · Yunfan Lu, Zipeng Wang, Minjie Liu, Hongjian Wang, Lin Wang

Event cameras sense the intensity changes asynchronously and produce event streams with high dynamic range and low latency. This has inspired research endeavors utilizing events to guide the challenging video superresolution (VSR) task. In this paper, we make the first attempt to address a novel problem of achieving VSR at random scales by taking advantages of the high temporal resolution property of events. This is hampered by the difficulties of representing the spatial-temporal information of events when guiding VSR. To this end, we propose a novel framework that incorporates the spatial-temporal interpolation of events to VSR in a unified framework. Our key idea is to learn implicit neural representations from queried spatial-temporal coordinates and features from both RGB frames and events. Our method contains three parts. Specifically, the Spatial-Temporal Fusion (STF) module first learns the 3D features from events and RGB frames. Then, the Temporal Filter (TF) module unlocks more explicit motion information from the events near the queried timestamp and generates the 2D features. Lastly, the SpatialTemporal Implicit Representation (STIR) module recovers the SR frame in arbitrary resolutions from the outputs of these two modules. In addition, we collect a real-world dataset with spatially aligned events and RGB frames. Extensive experiments show that our method significantly surpasses the prior-arts and achieves VSR with random scales, e.g., 6.5. Code and dataset are available at https: //vlis2022.github.io/cvpr23/egvsr.

📄 PDF Abstract BibTeX arXiv:2303.13767

Code (1)

yunfanLu/INR-Event-VSR 공식 구현 pytorch

Tasks

Super-ResolutionVideo Super-Resolution

Similar Papers 제목 키워드 기반

TVTA: Trajectory-Aware Viseme-Guided Temporal Aggregation for Event-Based Lip Reading

2026-07-09 · Jingrong Zheng, Hongwei Ren, Xiangqian Wu arxiv

Event-based lip reading has recently emerged as a promising direction for visual speech recognition, benefiting from the high temporal resolution and motion sensitivity of event cameras. However, existing methods typical…

Visual Speech RecognitionLip Reading

Spatial-Temporal Meta-path Guided Explainable Crime Prediction

2022-05-04 · Yuting Sun, Tong Chen, Hongzhi Yin

Exposure to crime and violence can harm individuals' quality of life and the economic growth of communities. In light of the rapid development in machine learning, there is a rise in the need to explore automated solutio…

BIG-bench Machine LearningCrime PredictionPrediction

UniINR: Event-guided Unified Rolling Shutter Correction, Deblurring, and Interpolation

2023-05-24 · Yunfan Lu, Guoqiang Liang, Yusheng Wang, Lin Wang 외

Video frames captured by rolling shutter (RS) cameras during fast camera movement frequently exhibit RS distortion and blur simultaneously. Naturally, recovering high-frame-rate global shutter (GS) sharp frames from an R…

DeblurringImage RestorationRolling Shutter Correction

Subspace Implicit Neural Representations for Real-Time Cardiac Cine MR Imaging

2024-12-17 · Wenqi Huang, Veronika Spieker, Siying Xu, Gastao Cruz 외

Conventional cardiac cine MRI methods rely on retrospective gating, which limits temporal resolution and the ability to capture continuous cardiac dynamics, particularly in patients with arrhythmias and beat-to-beat vari…

Diagnostic

EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events

2025-05-07 · CVPR 2025 1 · Shuoyan Wei, Feng Li, Shengeng Tang, Yao Zhao 외

Continuous space-time video super-resolution (C-STVSR) endeavors to upscale videos simultaneously at arbitrary spatial and temporal scales, which has recently garnered increasing interest. However, prevailing methods str…

Space-time Video Super-resolutionSuper-ResolutionVideo Super-Resolution