paper-with-me

홈 › Papers

Video Frame Interpolation for Polarization via Swin-Transformer

2024-06-17 · Feng Huang, Xin Zhang, YiXuan Xu, Xuesong Wang, Xianyu Wu

Video Frame Interpolation (VFI) has been extensively explored and demonstrated, yet its application to polarization remains largely unexplored. Due to the selective transmission of light by polarized filters, longer exposure times are typically required to ensure sufficient light intensity, which consequently lower the temporal sample rates. Furthermore, because polarization reflected by objects varies with shooting perspective, focusing solely on estimating pixel displacement is insufficient to accurately reconstruct the intermediate polarization. To tackle these challenges, this study proposes a multi-stage and multi-scale network called Swin-VFI based on the Swin-Transformer and introduces a tailored loss function to facilitate the network's understanding of polarization changes. To ensure the practicality of our proposed method, this study evaluates its interpolated frames in Shape from Polarization (SfP) and Human Shape Reconstruction tasks, comparing them with other state-of-the-art methods such as CAIN, FLAVR, and VFIT. Experimental results demonstrate our approach's superior reconstruction accuracy across all tasks.

📄 PDF Abstract BibTeX arXiv:2406.11371

Code (0)

등록된 구현이 없습니다.

Tasks

Video Frame Interpolation

Methods 이 논문이 사용한 방법론

FLAVR 설명 없음

Similar Papers 제목 키워드 기반

A Perceptual Quality Metric for Video Frame Interpolation

2022-10-04 · Qiqi Hou, Abhijay Ghildyal, Feng Liu

Research on video frame interpolation has made significant progress in recent years. However, existing methods mostly use off-the-shelf metrics to measure the quality of interpolation results with the exception of a few …

Video Frame Interpolation

SwinTExCo: Exemplar-based video colorization using Swin Transformer

2025-01-15 · Expert Systems with Applications 2025 1 · Duong Thanh Tran, Nguyen Doan Hieu Nguyen, Trung Thanh Pham, Phuong-Nam Tran 외

Video colorization represents a compelling domain within the field of Computer Vision. The traditional approach in this field relies on Convolutional Neural Networks (CNNs) to extract features from each video frame and e…

ColorizationVideo Restoration

Blur Interpolation Transformer for Real-World Motion from Blur

2022-11-21 · CVPR 2023 1 · Zhihang Zhong, Mingdeng Cao, Xiang Ji, Yinqiang Zheng 외

This paper studies the challenging problem of recovering motion from blur, also known as joint deblurring and interpolation or blur temporal super-resolution. The challenges are twofold: 1) the current methods still leav…

DeblurringSuper-Resolution

SwinBERT: End-to-End Transformers with Sparse Attention for Video Captioning

2021-11-25 · CVPR 2022 1 · Kevin Lin, Linjie Li, Chung-Ching Lin, Faisal Ahmed 외

The canonical approach to video captioning dictates a caption generation model to learn from offline-extracted dense video features. These feature extractors usually operate on video frames sampled at a fixed frame rate …

Caption GenerationQuestion AnsweringVideo CaptioningVideo Question Answering+1

Video Swin Transformer

2021-06-24 · CVPR 2022 1 · Ze Liu, Jia Ning, Yue Cao, Yixuan Wei 외

The vision community is witnessing a modeling shift from CNNs to Transformers, where pure Transformer architectures have attained top accuracy on the major video recognition benchmarks. These video models are all built o…

Action ClassificationAction RecognitionGeneral ClassificationInductive Bias+3