paper-with-me

Papers

Multi-Modal Domain Adaptation Across Video Scenes for Temporal Video Grounding

2023-12-21 · Haifeng Huang, Yang Zhao, Zehan Wang, Yan Xia, Zhou Zhao

Temporal Video Grounding (TVG) aims to localize the temporal boundary of a specific segment in an untrimmed video based on a given language query. Since datasets in this domain are often gathered from limited video scenes, models tend to overfit to scene-specific factors, which leads to suboptimal performance when encountering new scenes in real-world applications. In a new scene, the fine-grained annotations are often insufficient due to the expensive labor cost, while the coarse-grained video-query pairs are easier to obtain. Thus, to address this issue and enhance model performance on new scenes, we explore the TVG task in an unsupervised domain adaptation (UDA) setting across scenes for the first time, where the video-query pairs in the source scene (domain) are labeled with temporal boundaries, while those in the target scene are not. Under the UDA setting, we introduce a novel Adversarial Multi-modal Domain Adaptation (AMDA) method to adaptively adjust the model's scene-related knowledge by incorporating insights from the target data. Specifically, we tackle the domain gap by utilizing domain discriminators, which help identify valuable scene-related features effective across both domains. Concurrently, we mitigate the semantic gap between different modalities by aligning video-query pairs with related semantics. Furthermore, we employ a mask-reconstruction approach to enhance the understanding of temporal semantics within a scene. Extensive experiments on Charades-STA, ActivityNet Captions, and YouCook2 demonstrate the effectiveness of our proposed method.

📄 PDF Abstract BibTeX arXiv:2312.13633

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationUnsupervised Domain AdaptationVideo Grounding

Similar Papers 제목 키워드 기반

Learning Cross-modal Contrastive Features for Video Domain Adaptation

2021-08-26 · ICCV 2021 10 · Donghyun Kim, Yi-Hsuan Tsai, Bingbing Zhuang, Xiang Yu 외

Learning transferable and domain adaptive feature representations from videos is important for video-relevant tasks such as action recognition. Existing video domain adaptation methods mainly rely on adversarial feature …

Action RecognitionContrastive LearningDomain AdaptationOptical Flow Estimation

Cross-View Cross-Modal Unsupervised Domain Adaptation for Driver Monitoring System

2025-11-15 · Aditi Bhalla, Christian Hellert, Enkelejda Kasneci arxiv

Driver distraction remains a leading cause of road traffic accidents, contributing to thousands of fatalities annually across the globe. While deep learning-based driver activity recognition methods have shown promise in…

Unsupervised Domain AdaptationContrastive LearningActivity Recognition

Modality-Collaborative Low-Rank Decomposers for Few-Shot Video Domain Adaptation

2025-11-24 · Yuyang Wanyan, Xiaoshan Yang, Weiming Dong, Changsheng Xu arxiv

In this paper, we study the challenging task of Few-Shot Video Domain Adaptation (FSVDA). The multimodal nature of videos introduces unique challenges, necessitating the simultaneous consideration of both domain alignmen…

Domain Adaptation

Source-free Video Domain Adaptation by Learning Temporal Consistency for Action Recognition

2022-03-09 · Yuecong Xu, Jianfei Yang, Haozhi Cao, Keyu Wu 외

Video-based Unsupervised Domain Adaptation (VUDA) methods improve the robustness of video models, enabling them to be applied to action recognition tasks across different environments. However, these methods require cons…

Action RecognitionDomain AdaptationSource-Free Domain AdaptationUnsupervised Domain Adaptation

Modality-Collaborative Test-Time Adaptation for Action Recognition

2024-01-01 · CVPR 2024 1 · Baochen Xiong, Xiaoshan Yang, Yaguang Song, YaoWei Wang 외

Video-based Unsupervised Domain Adaptation (VUDA) method improves the generalization of the video model enabling it to be applied to action recognition tasks in different environments. However these methods require c…

Action RecognitionDomain AdaptationTest-time AdaptationUnsupervised Domain Adaptation