paper-with-me

홈 › Papers

Real-World Video for Zoom Enhancement based on Spatio-Temporal Coupling

2023-06-24 · Zhiling Guo, Yinqiang Zheng, Haoran Zhang, Xiaodan Shi, Zekun Cai, Ryosuke Shibasaki, Jinyue Yan

In recent years, single-frame image super-resolution (SR) has become more realistic by considering the zooming effect and using real-world short- and long-focus image pairs. In this paper, we further investigate the feasibility of applying realistic multi-frame clips to enhance zoom quality via spatio-temporal information coupling. Specifically, we first built a real-world video benchmark, VideoRAW, by a synchronized co-axis optical system. The dataset contains paired short-focus raw and long-focus sRGB videos of different dynamic scenes. Based on VideoRAW, we then presented a Spatio-Temporal Coupling Loss, termed as STCL. The proposed STCL is intended for better utilization of information from paired and adjacent frames to align and fuse features both temporally and spatially at the feature level. The outperformed experimental results obtained in different zoom scenarios demonstrate the superiority of integrating real-world video dataset and STCL into existing SR models for zoom quality enhancement, and reveal that the proposed method can serve as an advanced and viable tool for video zoom.

📄 PDF Abstract BibTeX arXiv:2306.13875

Code (0)

등록된 구현이 없습니다.

Tasks

Image Super-ResolutionSuper-Resolution

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Real-time Surgical Environment Enhancement for Robot-Assisted Minimally Invasive Surgery Based on Super-Resolution

2020-11-08 · Ruoxi Wang, Dandan Zhang, QingBiao Li, Xiao-Yun Zhou 외

In Robot-Assisted Minimally Invasive Surgery (RAMIS), a camera assistant is normally required to control the position and zooming ratio of the laparoscope, following the surgeon's instructions. However, moving the laparo…

Depth EstimationGenerative Adversarial NetworkSuper-ResolutionVideo Super-Resolution

ZoomNeXt: A Unified Collaborative Pyramid Network for Camouflaged Object Detection

2023-10-31 · Youwei Pang, Xiaoqi Zhao, Tian-Zhu Xiang, Lihe Zhang 외

Recent camouflaged object detection (COD) attempts to segment objects visually blended into their surroundings, which is extremely complex and difficult in real-world scenarios. Apart from the high intrinsic similarity b…

Camouflaged Object Segmentation

Zoom-VQA: Patches, Frames and Clips Integration for Video Quality Assessment

2023-04-13 · Kai Zhao, Kun Yuan, Ming Sun, Xing Wen

Video quality assessment (VQA) aims to simulate the human perception of video quality, which is influenced by factors ranging from low-level color and texture details to high-level semantic content. To effectively model …

Video Quality AssessmentVisual Question Answering (VQA)

WonderZoom: Multi-Scale 3D World Generation

2025-12-09 · Jin Cao, Hong-Xing Yu, Jiajun Wu arxiv

We present WonderZoom, a novel approach to generating 3D scenes with contents across multiple spatial scales from a single image. Existing 3D world generation models remain limited to single-scale synthesis and cannot pr…

ReBotNet: Fast Real-time Video Enhancement

2023-03-23 · Jeya Maria Jose Valanarasu, Rahul Garg, Andeep Toor, Xin Tong 외

Most video restoration networks are slow, have high computational load, and can't be used for real-time video enhancement. In this work, we design an efficient and fast framework to perform real-time video enhancement fo…

DecoderVideo EnhancementVideo Restoration