paper-with-me

홈 › Papers

Leveraging Content and Context Cues for Low-Light Image Enhancement

2024-12-10 · Igor Morawski, Kai He, Shusil Dangi, Winston H. Hsu

Low-light conditions have an adverse impact on machine cognition, limiting the performance of computer vision systems in real life. Since low-light data is limited and difficult to annotate, we focus on image processing to enhance low-light images and improve the performance of any downstream task model, instead of fine-tuning each of the models which can be prohibitively expensive. We propose to improve the existing zero-reference low-light enhancement by leveraging the CLIP model to capture image prior and for semantic guidance. Specifically, we propose a data augmentation strategy to learn an image prior via prompt learning, based on image sampling, to learn the image prior without any need for paired or unpaired normal-light data. Next, we propose a semantic guidance strategy that maximally takes advantage of existing low-light annotation by introducing both content and context cues about the image training patches. We experimentally show, in a qualitative study, that the proposed prior and semantic guidance help to improve the overall image contrast and hue, as well as improve background-foreground discrimination, resulting in reduced over-saturation and noise over-amplification, common in related zero-reference methods. As we target machine cognition, rather than rely on assuming the correlation between human perception and downstream task performance, we conduct and present an ablation study and comparison with related zero-reference methods in terms of task-based performance across many low-light datasets, including image classification, object and face detection, showing the effectiveness of our proposed method.

📄 PDF Abstract BibTeX arXiv:2412.07693

Code (1)

igor-morawski/tmm-sem 공식 구현 pytorch

Tasks

Data AugmentationFace Detectionimage-classificationImage ClassificationImage EnhancementLow-Light Image EnhancementPrompt Learning

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
Focus 설명 없음

Similar Papers 제목 키워드 기반

Contextual Outpainting With Object-Level Contrastive Learning

2022-01-01 · CVPR 2022 1 · Jiacheng Li, Chang Chen, Zhiwei Xiong

We study the problem of contextual outpainting, which aims to hallucinate the missing background contents based on the remaining foreground contents. Existing image outpainting methods focus on completing object shap…

Contrastive LearningImage OutpaintingObject

Enhancing Cross-Modal Contextual Congruence for Crowdfunding Success using Knowledge-infused Learning

2024-02-06 · Trilok Padhi, Ugur Kursuncu, Yaman Kumar, Valerie L. Shalin 외

The digital landscape continually evolves with multimodality, enriching the online experience for users. Creators and marketers aim to weave subtle contextual cues from various modalities into congruent content to engage…

Common Sense ReasoningKnowledge GraphsMarketing

AV-EmoDialog: Chat with Audio-Visual Users Leveraging Emotional Cues

2024-12-23 · Se Jin Park, Yeonju Kim, Hyeongseop Rha, Bella Godiva 외

In human communication, both verbal and non-verbal cues play a crucial role in conveying emotions, intentions, and meaning beyond words alone. These non-linguistic information, such as facial expressions, eye contact, vo…

Advancing Automated Speaking Assessment Leveraging Multifaceted Relevance and Grammar Information

2025-06-19 · Hao-Chien Lu, Jhen-Ke Lin, Hong-Yun Lin, Chung-Chun Wang 외

Current automated speaking assessment (ASA) systems for use in multi-aspect evaluations often fail to make full use of content relevance, overlooking image or exemplar cues, and employ superficial grammar analysis that l…

KCD: Knowledge Walks and Textual Cues Enhanced Political Perspective Detection in News Media

2022-04-08 · NAACL 2022 7 · Wenqian Zhang, Shangbin Feng, Zilong Chen, Zhenyu Lei 외

Political perspective detection has become an increasingly important task that can help combat echo chambers and political polarization. Previous approaches generally focus on leveraging textual content to identify stanc…

ArticlesKnowledge GraphsRepresentation Learning