paper-with-me

홈 › Papers

Learning a Domain-Agnostic Visual Representation for Autonomous Driving via Contrastive Loss

2021-03-10 · Dongseok Shim, H. Jin Kim

Deep neural networks have been widely studied in autonomous driving applications such as semantic segmentation or depth estimation. However, training a neural network in a supervised manner requires a large amount of annotated labels which are expensive and time-consuming to collect. Recent studies leverage synthetic data collected from a virtual environment which are much easier to acquire and more accurate compared to data from the real world, but they usually suffer from poor generalization due to the inherent domain shift problem. In this paper, we propose a Domain-Agnostic Contrastive Learning (DACL) which is a two-stage unsupervised domain adaptation framework with cyclic adversarial training and contrastive loss. DACL leads the neural network to learn domain-agnostic representation to overcome performance degradation when there exists a difference between training and test data distribution. Our proposed approach achieves better performance in the monocular depth estimation task compared to previous state-of-the-art methods and also shows effectiveness in the semantic segmentation task.

📄 PDF Abstract BibTeX arXiv:2103.05902

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingContrastive LearningDepth EstimationDomain AdaptationMonocular Depth EstimationSegmentationSemantic SegmentationUnsupervised Domain Adaptation

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Commonsense Visual Sensemaking for Autonomous Driving: On Generalised Neurosymbolic Online Abduction Integrating Vision and Semantics

2020-12-28 · Jakob Suchan, Mehul Bhatt, Srikrishna Varadarajan

We demonstrate the need and potential of systematically integrated vision and semantics solutions for visual sensemaking in the backdrop of autonomous driving. A general neurosymbolic method for online visual sensemaking…

Autonomous DrivingQuestion AnsweringSpatial Reasoning

GeoDrive-Bench: Benchmarking Region-Specific Multimodal Reasoning in Autonomous Driving

2026-06-01 · Yingzi Ma, Chaowei Xiao, Ming Jiang arxiv

Vision-language models (VLMs) for autonomous driving have shown promising performance, but their ability to handle region-specific traffic rules remains underexplored, raising uncertainties about their deployment across …

Multimodal ReasoningScene UnderstandingAutonomous Driving

Zero-Shot Cross-City Generalization in End-to-End Autonomous Driving: Self-Supervised versus Supervised Representations

2026-03-12 · Fatemeh Naeinian, Ali Hamza, Haoran Zhu, Anna Choromanska arxiv

End-to-end autonomous driving models are typically trained on multi-city datasets using supervised ImageNet-pretrained backbones, yet their ability to generalize to unseen cities remains largely unexamined. When training…

Representation LearningAutonomous Driving

Vision and Language: Novel Representations and Artificial intelligence for Driving Scene Safety Assessment and Autonomous Vehicle Planning

2026-02-07 · Ross Greer, Maitrayee Keskar, Angel Martinez-Sanchez, Parthib Roy 외 arxiv

Vision-language models (VLMs) have recently emerged as powerful representation learning systems that align visual observations with natural language concepts, offering new opportunities for semantic reasoning in safety-c…

Visual Question AnsweringRepresentation LearningTrajectory PlanningAutonomous Driving

Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving

2025-12-11 · Jiawei Yang, Ziyu Chen, Yurong You, Yan Wang 외 arxiv

We present Flex, an efficient and effective scene encoder that addresses the computational bottleneck of processing high-volume multi-camera data in end-to-end autonomous driving. Flex employs a small set of learnable sc…

Autonomous Driving