paper-with-me

홈 › Papers

Toward Highly Efficient Semantic-Guided Machine Vision for Low-Light Object Detection

2024-12-20 · British Machine Vision Conference 2024 12 · Xin Feng, Junxian Zeng, Siping Wang, Zhenwei He

Detectors trained on well-lit data often experience significant performance degradation when applied to low-light conditions. To address this challenge, low-light enhancement methods are commonly employed to improve detection performance. However, existing human vision-oriented enhancement methods have shown limited effectiveness, which overlooks the semantic information for detection and achieves high computation costs. To overcome these limitations, we introduce a machine vision-oriented highly efficient low-light object detection method with the Efficient semantic-guided Machine Vision-oriented module (EMV). EMV can dynamically adapt to the object detection part based on end-to-end training and emphasize the semantic information for the detection. Besides, by lightening the network for feature decomposition and generating the enhanced image on latent space, EMV is a highly lightweight network for image enhancement, which contains only 27K parameters and achieves high inference speed. Extensive experiments conducted on ExDark and DarkFace datasets demonstrate that our method significantly improves detector performance in low-light environments.

📄 PDF Abstract BibTeX

Code (1)

Zeng555/EMV-YOLO pytorch

Tasks

2D Object DetectionImage Enhancementobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Lightweight Prompt-Guided CLIP Adaptation for Monocular Depth Estimation

2026-04-01 · Reyhaneh Ahani Manghotay, Jie Liang arxiv

Leveraging the rich semantic features of vision-language models (VLMs) like CLIP for monocular depth estimation tasks is a promising direction, yet often requires extensive fine-tuning or lacks geometric precision. We pr…

Monocular Depth Estimation

Video-guided Machine Translation with Global Video Context

2026-04-08 · Jian Chen, JinZe Lv, Zi Long, XiangHua Fu arxiv

Video-guided Multimodal Translation (VMT) has advanced significantly in recent years. However, most existing methods rely on locally aligned video segments paired one-to-one with subtitles, limiting their ability to capt…

Machine Translation

Semantic-Guided Zero-Shot Learning for Low-Light Image/Video Enhancement

2021-10-03 · Shen Zheng, Gaurav Gupta

Low-light images challenge both human perceptions and computer vision algorithms. It is crucial to make algorithms robust to enlighten low-light images for computational photography and computer vision applications such …

Image EnhancementLow-Light Image EnhancementSegmentationSemantic Segmentation+3

See&Say: Vision Language Guided Safe Zone Detection for Autonomous Package Delivery Drones

2026-04-14 · Mahyar Ghazanfari, Peng Wei arxiv

Autonomous drone delivery systems are rapidly advancing, but ensuring safe and reliable package drop-offs remains highly challenging in cluttered urban and suburban environments where accurately identifying suitable pack…

Semantic Segmentation

Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

2026-04-10 · Zibin Geng, Xuefeng Jiang, Jia Li, Zheng Li 외 arxiv

Prompt learning is a parameter-efficient approach for vision-language models, yet its robustness under label noise is less investigated. Visual content contains richer and more reliable semantic information, which remain…