paper-with-me

홈 › Papers

TimeRewind: Rewinding Time with Image-and-Events Video Diffusion

2024-03-20 · Jingxi Chen, Brandon Y. Feng, Haoming Cai, Mingyang Xie, Christopher Metzler, Cornelia Fermuller, Yiannis Aloimonos

This paper addresses the novel challenge of `rewinding'' time from a single captured image to recover the fleeting moments missed just before the shutter button is pressed. This problem poses a significant challenge in computer vision and computational photography, as it requires predicting plausible pre-capture motion from a single static frame, an inherently ill-posed task due to the high degree of freedom in potential pixel movements. We overcome this challenge by leveraging the emerging technology of neuromorphic event cameras, which capture motion information with high temporal resolution, and integrating this data with advanced image-to-video diffusion models. Our proposed framework introduces an event motion adaptor conditioned on event camera data, guiding the diffusion model to generate videos that are visually coherent and physically grounded in the captured events. Through extensive experimentation, we demonstrate the capability of our approach to synthesize high-quality videos that effectively `rewind'' time, showcasing the potential of combining event camera technology with generative models. Our work opens new avenues for research at the intersection of computer vision, computational photography, and generative modeling, offering a forward-thinking solution to capturing missed moments and enhancing future consumer cameras and smartphones. Please see the project page at https://timerewind.github.io/ for video results and code release.

📄 PDF Abstract BibTeX arXiv:2403.13800

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Scalable and Explainable Learner-Video Interaction Prediction using Multimodal Large Language Models

2026-04-06 · Dominik Glandorf, Fares Fawzi, Tanja Käser arxiv

Learners' use of video controls in educational videos provides implicit signals of cognitive processing and instructional design quality, yet the lack of scalable and explainable predictive models limits instructors' abi…

Reproducibility Study: Comparing Rewinding and Fine-tuning in Neural Network Pruning

2021-09-20 · Szymon Mikler

Scope of reproducibility: We are reproducing Comparing Rewinding and Fine-tuning in Neural Networks from arXiv:2003.02389. In this work the authors compare three different approaches to retraining neural networks after p…

Network Pruning

Comparing Rewinding and Fine-tuning in Neural Network Pruning

2020-03-05 · ICLR 2020 1 · Alex Renda, Jonathan Frankle, Michael Carbin

Many neural network pruning algorithms proceed in three steps: train the network to completion, remove unwanted structure to compress the network, and retrain the remaining structure to recover lost accuracy. The standar…

Network Pruning

Network Pruning That Matters: A Case Study on Retraining Variants

2021-05-07 · ICLR 2021 1 · Duong H. Le, Binh-Son Hua

Network pruning is an effective method to reduce the computational expense of over-parameterized neural networks for deployment on low-resource systems. Recent state-of-the-art techniques for retraining pruned networks s…

Network PruningSentence

Restage4D: Reanimating Deformable 3D Reconstruction from a Single Video

2025-08-08 · Jixuan He, Chieh Hubert Lin, Lu Qi, Ming-Hsuan Yang arxiv

Creating deformable 3D content has gained increasing attention with the rise of text-to-image and image-to-video generative models. While these models provide rich semantic priors for appearance, they struggle to capture…

3D Reconstruction