paper-with-me

홈 › Papers

D3T: Distinctive Dual-Domain Teacher Zigzagging Across RGB-Thermal Gap for Domain-Adaptive Object Detection

2024-03-14 · CVPR 2024 1 · Dinh Phat Do, TaeHoon Kim, Jaemin Na, Jiwon Kim, Keonho Lee, Kyunghwan Cho, Wonjun Hwang

Domain adaptation for object detection typically entails transferring knowledge from one visible domain to another visible domain. However, there are limited studies on adapting from the visible to the thermal domain, because the domain gap between the visible and thermal domains is much larger than expected, and traditional domain adaptation can not successfully facilitate learning in this situation. To overcome this challenge, we propose a Distinctive Dual-Domain Teacher (D3T) framework that employs distinct training paradigms for each domain. Specifically, we segregate the source and target training sets for building dual-teachers and successively deploy exponential moving average to the student model to individual teachers of each domain. The framework further incorporates a zigzag learning method between dual teachers, facilitating a gradual transition from the visible to thermal domains during training. We validate the superiority of our method through newly designed experimental protocols with well-known thermal datasets, i.e., FLIR and KAIST. Source code is available at https://github.com/EdwardDo69/D3T .

📄 PDF Abstract BibTeX arXiv:2403.09359

Code (1)

edwarddo69/d3t 공식 구현 pytorch

Tasks

Domain Adaptationobject-detectionObject Detection

Similar Papers 제목 키워드 기반

AM-RADIO: Agglomerative Vision Foundation Model -- Reduce All Domains Into One

2023-12-10 · Mike Ranzinger, Greg Heinrich, Jan Kautz, Pavlo Molchanov

A handful of visual foundation models (VFMs) have recently emerged as the backbones for numerous downstream tasks. VFMs like CLIP, DINOv2, SAM are trained with distinct objectives, exhibiting unique characteristics for v…

AllBenchmarkingobject-detectionObject Detection+1

AM-RADIO: Agglomerative Vision Foundation Model Reduce All Domains Into One

2024-01-01 · CVPR 2024 1 · Mike Ranzinger, Greg Heinrich, Jan Kautz, Pavlo Molchanov

A handful of visual foundation models (VFMs) have recently emerged as the backbones for numerous downstream tasks. VFMs like CLIP DINOv2 SAM are trained with distinct objectives exhibiting unique characteristics for …

AllBenchmarkingobject-detectionObject Detection+1

TimeBalance: Temporally-Invariant and Temporally-Distinctive Video Representations for Semi-Supervised Action Recognition

2023-03-28 · CVPR 2023 1 · Ishan Rajendrakumar Dave, Mamshad Nayeem Rizve, Chen Chen, Mubarak Shah

Semi-Supervised Learning can be more beneficial for the video domain compared to images because of its higher annotation cost and dimensionality. Besides, any video understanding task requires reasoning over both spatial…

Action RecognitionOptical Flow EstimationVideo Understanding

Unbiased Mean Teacher for Cross-domain Object Detection

2020-03-02 · CVPR 2021 1 · Jinhong Deng, Wen Li, Yu-Hua Chen, Lixin Duan

Cross-domain object detection is challenging, because object detection model is often vulnerable to data variance, especially to the considerable domain shift between two distinctive domains. In this paper, we propose a …

Objectobject-detectionObject DetectionSmall Data Image Classification+1

Recognizability of Individual Creative Style Within and Across Domains: Preliminary Studies

2010-05-10 · Liane Gabora

It is hypothesized that creativity arises from the self-mending capacity of an internal model of the world, or worldview. The uniquely honed worldview of a creative individual results in a distinctive style that is recog…