paper-with-me

Papers

DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations

2025-08-14 · Hang Jin, Chenqiang Gao, Junjie Guo, Fangcen Liu, Kanghui Tian, Qinyao Chang arxiv

Infrared-visible object detection has shown great potential in real-world applications, enabling robust all-day perception by leveraging the complementary information of infrared and visible images. However, existing methods typically require dual-modality annotations to output decoupled detection results, leading to high annotation costs and limiting scalability in large-scale remote sensing applications. To address this challenge, we propose a novel infrared-visible \textbf{D}ecoupled \textbf{O}bject \textbf{D}etection framework with \textbf{S}ingle-modality \textbf{A}nnotations, called DOD-SA. It is built upon a Single- and Dual-Modality Collaborative Teacher-Student Network (CoSD-TSNet), which consists of a single-modality branch (SM-Branch) and a dual-modality decoupled branch (DMD-Branch). This design enables cross-modality knowledge transfer from the labeled modality to the unlabeled modality, and facilitates effective cross-branch supervision. To further improve the quality of pseudo-labels, we introduce a Progressive and Self-Tuning Training Strategy (PaST) that trains the model in three stages: 1) SM-Branch self-training, 2) SM-Branch guiding the learning of DMD-Branch, and 3) DMD-Branch refinement. In addition, we design a Pseudo Label Assigner (PLA) to match labels across modalities, explicitly addressing modality misalignment during training.

📄 PDF Abstract BibTeX arXiv:2508.10445

Code (0)

등록된 구현이 없습니다.

Tasks

Object Detection

Similar Papers 제목 키워드 기반

DPDETR: Decoupled Position Detection Transformer for Infrared-Visible Object Detection

2024-08-12 · Junjie Guo, Chenqiang Gao, Fangcen Liu, Deyu Meng

Infrared-visible object detection aims to achieve robust object detection by leveraging the complementary information of infrared and visible image pairs. However, the commonly existing modality misalignment problem pres…

DecoderObjectobject-detectionObject Detection+2

DEYOLO: Dual-Feature-Enhancement YOLO for Cross-Modality Object Detection

2024-12-06 · Yishuo Chen, Boran Wang, Xinyu Guo, Wenbin Zhu 외

Object detection in poor-illumination environments is a challenging task as objects are usually not clearly visible in RGB images. As infrared images provide additional clear edge information that complements RGB images,…

Objectobject-detectionObject Detection

MetaFusion: Infrared and Visible Image Fusion via Meta-Feature Embedding From Object Detection

2023-01-01 · CVPR 2023 1 · Wenda Zhao, Shigeng Xie, Fan Zhao, You He 외

Fusing infrared and visible images can provide more texture details for subsequent object detection task. Conversely, detection task furnishes object semantic information to improve the infrared and visible image fus…

Infrared And Visible Image FusionMeta-LearningObjectobject-detection+1

IV-tuning: Parameter-Efficient Transfer Learning for Infrared-Visible Tasks

2024-12-21 · Yaming Zhang, Chenqiang Gao, Fangcen Liu, Junjie Guo 외

Infrared-visible (IR-VIS) tasks, such as semantic segmentation and object detection, greatly benefit from the advantage of combining infrared and visible modalities. To inherit the general representations of the Vision F…

object-detectionObject DetectionSemantic SegmentationTransfer Learning

CDUPatch: Color-Driven Universal Adversarial Patch Attack for Dual-Modal Visible-Infrared Detectors

2025-04-15 · Jiahuan Long, Wen Yao, Tingsong Jiang, Chao Ma

Adversarial patches are widely used to evaluate the robustness of object detection systems in real-world scenarios. These patches were initially designed to deceive single-modal detectors (e.g., visible or infrared) and …

Data Augmentationobject-detectionObject Detection