Factors of Influence for Transfer Learning across Diverse Appearance Domains and Task Types
Transfer learning enables to re-use knowledge learned on a source task to help learning a target task. A simple form of transfer learning is common in current state-of-the-art computer vision models, i.e. pre-training a model for image classification on the ILSVRC dataset, and then fine-tune on any target task. However, previous systematic studies of transfer learning have been limited and the circumstances in which it is expected to work are not fully understood. In this paper we carry out an extensive experimental exploration of transfer learning across vastly different image domains (consumer photos, autonomous driving, aerial imagery, underwater, indoor scenes, synthetic, close-ups) and task types (semantic segmentation, object detection, depth estimation, keypoint detection). Importantly, these are all complex, structured output tasks types relevant to modern computer vision applications. In total we carry out over 2000 transfer learning experiments, including many where the source and target come from different image domains, task types, or both. We systematically analyze these experiments to understand the impact of image domain, task type, and dataset size on transfer learning performance. Our study leads to several insights and concrete recommendations: (1) for most tasks there exists a source which significantly outperforms ILSVRC'12 pre-training; (2) the image domain is the most important factor for achieving positive transfer; (3) the source dataset should \emph{include} the image domain of the target dataset to achieve best results; (4) at the same time, we observe only small negative effects when the image domain of the source task is much broader than that of the target; (5) transfer across task types can be beneficial, but its success is heavily dependent on both the source and target task types.
Code (0)
등록된 구현이 없습니다.
Tasks
Autonomous DrivingDepth Estimationimage-classificationImage ClassificationKeypoint Detectionobject-detectionObject DetectionSemantic SegmentationTransfer LearningSimilar Papers 제목 키워드 기반
Personality Perception in Human Videos Altered by Motion Transfer Networks
The successful portrayal of personality in digital characters improves communication and immersion. Current research focuses on expressing personality through modifying animations using heuristic rules or data-driven mod…
Generating Furry Cars: Disentangling Object Shape & Appearance across Multiple Domains
We consider the novel task of learning disentangled representations of object shape and appearance across multiple domains (e.g., dogs and cars). The goal is to learn a generative model that learns an intermediate distri…
DisentanglementObjectGenerating Furry Cars: Disentangling Object Shape and Appearance across Multiple Domains
We consider the novel task of learning disentangled representations of object shape and appearance across multiple domains (e.g., dogs and cars). The goal is to learn a generative model that learns an intermediate distr…
DisentanglementObjectHOT: Harmonic-Constrained Optimal Transport for Remote Photoplethysmography Domain Adaptation
Remote photoplethysmography (rPPG) enables non-contact physiological measurement from facial videos; however, its practical deployment is often hindered by substantial performance degradation under domain shift. While re…
Domain AdaptationTowards Generalizable Multi-Object Tracking
Multi-Object Tracking MOT encompasses various tracking scenarios, each characterized by unique traits. Effective trackers should demonstrate a high degree of generalizability across diverse scenarios. However, existing t…
Domain GeneralizationMulti-Object TrackingObjectObject Tracking