paper-with-me

Papers

Supervised Image Translation from Visible to Infrared Domain for Object Detection

2024-08-03 · Prahlad Anand, Qiranul Saadiyean, Aniruddh Sikdar, Nalini N, Suresh Sundaram

This study aims to learn a translation from visible to infrared imagery, bridging the domain gap between the two modalities so as to improve accuracy on downstream tasks including object detection. Previous approaches attempt to perform bi-domain feature fusion through iterative optimization or end-to-end deep convolutional networks. However, we pose the problem as similar to that of image translation, adopting a two-stage training strategy with a Generative Adversarial Network and an object detection model. The translation model learns a conversion that preserves the structural detail of visible images while preserving the texture and other characteristics of infrared images. Images so generated are used to train standard object detection frameworks including Yolov5, Mask and Faster RCNN. We also investigate the usefulness of integrating a super-resolution step into our pipeline to further improve model accuracy, and achieve an improvement of as high as 5.3% mAP.

📄 PDF Abstract BibTeX arXiv:2408.01843

Code (0)

등록된 구현이 없습니다.

Tasks

Generative Adversarial NetworkObjectobject-detectionObject DetectionSuper-ResolutionTranslation

Similar Papers 제목 키워드 기반

Cross-Modal Spherical Aggregation for Weakly Supervised Remote Sensing Shadow Removal

2024-06-25 · Kaichen Chi, Wei Jing, Junjie Li, Qiang Li 외

Remote sensing shadow removal, which aims to recover contaminated surface information, is tricky since shadows typically display overwhelmingly low illumination intensities. In contrast, the infrared image is robust towa…

Shadow Removal

I2V-GAN: Unpaired Infrared-to-Visible Video Translation

2021-08-02 · Shuang Li, Bingfeng Han, Zhenjie Yu, Chi Harold Liu 외

Human vision is often adversely affected by complex environmental factors, especially in night vision scenarios. Thus, infrared cameras are often leveraged to help enhance the visual effects via detecting infrared radiat…

object-detectionObject DetectionTranslation

ROMA: Cross-Domain Region Similarity Matching for Unpaired Nighttime Infrared to Daytime Visible Video Translation

2022-04-26 · Zhenjie Yu, Kai Chen, Shuang Li, Bingfeng Han 외

Infrared cameras are often utilized to enhance the night vision since the visible light cameras exhibit inferior efficacy without sufficient illumination. However, infrared data possesses inadequate color contrast and re…

Translation

VI-Diff: Unpaired Visible-Infrared Translation Diffusion Model for Single Modality Labeled Visible-Infrared Person Re-identification

2023-10-06 · Han Huang, Yan Huang, Liang Wang

Visible-Infrared person re-identification (VI-ReID) in real-world scenarios poses a significant challenge due to the high cost of cross-modality data annotation. Different sensing cameras, such as RGB/IR cameras for good…

Image-to-Image TranslationPerson Re-IdentificationTranslation

CM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images

2025-03-12 · Bin Hu, Chenqiang Gao, Shurui Liu, Junjie Guo 외

The image translation method represents a crucial approach for mitigating information deficiencies in the infrared and visible modalities, while also facilitating the enhancement of modality-specific datasets. However, e…

Translation