paper-with-me

Papers

DST-Net: A Dual-Stream Transformer with Illumination-Independent Feature Guidance and Multi-Scale Spatial Convolution for Low-Light Image Enhancement

2026-03-17 · Yicui Shi, Yuhan Chen, Xiangfei Huang, Zhenguo Wang, Wenxuan Yu, Ying Fang arxiv

Low-light image enhancement aims to restore the visibility of images captured by visual sensors in dim environments by addressing their inherent signal degradations, such as luminance attenuation and structural corruption. Although numerous algorithms attempt to improve image quality, existing methods often cause a severe loss of intrinsic signal priors. To overcome these challenges, we propose a Dual-Stream Transformer Network (DST-Net) based on illumination-agnostic signal prior guidance and multi-scale spatial convolutions. First, to address the loss of critical signal features under low-light conditions, we design a feature extraction module. This module integrates Difference of Gaussians (DoG), LAB color space transformations, and VGG-16 for texture extraction, utilizing decoupled illumination-agnostic features as signal priors to continuously guide the enhancement process. Second, we construct a dual-stream interaction architecture. By employing a cross-modal attention mechanism, the network leverages the extracted priors to dynamically rectify the deteriorated signal representation of the enhanced image, ultimately achieving iterative enhancement through differentiable curve estimation. Furthermore, to overcome the inability of existing methods to preserve fine structures and textures, we propose a Multi-Scale Spatial Fusion Block (MSFB) featuring pseudo-3D and 3D gradient operator convolutions. This module integrates explicit gradient operators to recover high-frequency edges while capturing inter-channel spatial correlations via multi-scale spatial convolutions. Extensive evaluations and ablation studies demonstrate that DST-Net achieves superior performance in subjective visual quality and objective metrics. Specifically, our method achieves a PSNR of 25.64 dB on the LOL dataset. Subsequent validation on the LSRW dataset further confirms its robust cross-scene generalization.

📄 PDF Abstract BibTeX arXiv:2603.16482

Code (0)

등록된 구현이 없습니다.

Tasks

Low-Light Image Enhancement

Results from the Paper

RankTaskDatasetModelMetrics
#1 Low-Light Image Enhancement LSRW Dual-Stream Average PSNR: 25.64

Similar Papers 제목 키워드 기반

Wound Segmentation with Dynamic Illumination Correction and Dual-view Semantic Fusion

2022-07-12 · Honghui Liu, Changjian Wang, Kele Xu, Fangzhao Li 외

Wound image segmentation is a critical component for the clinical diagnosis and in-time treatment of wounds. Recently, deep learning has become the mainstream methodology for wound image segmentation. However, the pre-pr…

Image SegmentationSegmentationSemantic Segmentation

ISALux: Illumination and Segmentation Aware Transformer Employing Mixture of Experts for Low Light Image Enhancement

2025-08-25 · Raul Balmez, Alexandru Brateanu, Ciprian Orhei, Codruta Ancuti 외 arxiv

We introduce ISALux, a novel transformer-based approach for Low-Light Image Enhancement (LLIE) that seamlessly integrates illumination and semantic priors. Our architecture includes an original self-attention block, Hybr…

Low-Light Image EnhancementSemantic Segmentation

The Dual-Stream Transformer: Channelized Architecture for Interpretable Language Modeling

2026-03-08 · J. Clayton Kerce, Alexis Fox arxiv

Standard transformers entangle all computation in a single residual stream, obscuring which components perform which functions. We introduce the Dual-Stream Transformer, which decomposes the residual stream into two func…

Retinex-guided Histogram Transformer for Mask-free Shadow Removal

2025-04-18 · Wei Dong, Han Zhou, Seyed Amirreza Mousavi, Jun Chen

While deep learning methods have achieved notable progress in shadow removal, many existing approaches rely on shadow masks that are difficult to obtain, limiting their generalization to real-world scenes. In this work, …

Shadow Removal

DRPFNet: Dual-domain Residual Progressive Fusion Network for RGB-Thermal Object Detection

2026-08-04 · Zian Wang, Changchun Li arxiv

RGB-thermal (RGB-T) object detection aims to fuse complementary information from visible and thermal modalities to achieve robust detection under varying illumination and weather conditions. Current methods typically emp…

Object LocalizationObject Detection