paper-with-me

홈 › Papers

Dynamic Weight-based Temporal Aggregation for Low-light Video Enhancement Under Extreme Noise

2025-10-10 · Ruirui Lin, Guoxi Huang, Nantheera Anantrasirichai arxiv

Low-light video enhancement (LLVE) is challenging due to noise, low contrast, and color degradation. While learning-based methods enable fast inference, they often fail under heavy real-world noise because they do not sufficiently exploit long-term temporal cues. We propose DWTA-Net, a novel deep-learning recurrent LLVE framework with a recurrent design. DWTA-Net adopts an integrated two-stage architecture: Stage I restores local structure and color via multi-frame alignment for temporally consistent Mamba-based enhancement, while Stage II performs recurrent refinement using a novel dynamic weight-based temporal aggregation guided by optical flow, functioning as a recurrent denoiser that adapts to motion. We further introduce a texture-adaptive loss that preserves fine details in textured regions while suppressing noise in homogeneous areas. Experiments on real-world low-light footage show that DWTA-Net achieves stronger noise suppression and fewer artifacts, delivering superior visual quality compared with state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2510.09450

Code (0)

등록된 구현이 없습니다.

Tasks

Video Enhancement

Similar Papers 제목 키워드 기반

Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection

2024-12-06 · Khurram Azeem Hashmi, Talha Uddin Sheikh, Didier Stricker, Muhammad Zeshan Afzal

The primary challenge in Video Object Detection (VOD) is effectively exploiting temporal information to enhance object representations. Traditional strategies, such as aggregating region proposals, often suffer from feat…

GPUMulti-Object TrackingObjectobject-detection+4

TAM: Temporal Adaptive Module for Video Recognition

2020-05-14 · ICCV 2021 10 · Zhao-Yang Liu, Li-Min Wang, Wayne Wu, Chen Qian 외

Video data is with complex temporal dynamics due to various factors such as camera motion, speed variation, and different activities. To effectively capture this diverse motion pattern, this paper presents a new temporal…

Action RecognitionVideo Recognition

The Devil is in Temporal Token: High Quality Video Reasoning Segmentation

2025-01-15 · CVPR 2025 1 · Sitong Gong, Yunzhi Zhuge, Lu Zhang, Zongxin Yang 외

Existing methods for Video Reasoning Segmentation rely heavily on a single special token to represent the object in the keyframe or the entire video, inadequately capturing spatial complexity and inter-frame motion. To o…

Reasoning SegmentationReferring Expression SegmentationReferring Video Object SegmentationSegmentation

Deep Video Matting via Spatio-Temporal Alignment and Aggregation

2021-04-22 · CVPR 2021 1 · Yanan sun, Guanzhi Wang, Qiao Gu, Chi-Keung Tang 외

Despite the significant progress made by deep learning in natural image matting, there has been so far no representative work on deep learning for video matting due to the inherent technical challenges in reasoning tempo…

DecoderDeep LearningImage MattingOptical Flow Estimation+1

Spatio-Temporal Similarity Volume Aggregation for Open-Vocabulary Action Recognition

2026-05-22 · Yerim So, Jiyeong Kim, Jiwon Yoon, Dongbo Min arxiv

Recent Open-Vocabulary Action Recognition (OVAR) methods typically aggregate visual features into a global representation before computing text alignment, a process that obscures local patch information and fine-grained …

Action Recognition