paper-with-me

Papers

Modality-Collaborative Test-Time Adaptation for Action Recognition

2024-01-01 · CVPR 2024 1 · Baochen Xiong, Xiaoshan Yang, Yaguang Song, YaoWei Wang, Changsheng Xu

Video-based Unsupervised Domain Adaptation (VUDA) method improves the generalization of the video model enabling it to be applied to action recognition tasks in different environments. However these methods require continuous access to source data during the adaptation process which are impractical in real scenarios where the source videos are not available with concerns in transmission efficiency or privacy issues. To address this problem in this paper we propose to solve the Multimodal Video Test-Time Adaptation task (MVTTA). Existing image-based TTA methods cannot be directly applied to this task because video have domain shift in multimodal and temporal which brings difficulties to adaptation. To address the above challenges we propose a Modality-Collaborative Test-Time Adaptation (MC-TTA) Network. We maintain teacher and student memory banks respectively for generating pseudo-prototypes and target-prototypes. In the teacher model we propose Self-assembled Source-friendly Feature Reconstruction (SSFR) module to encourage the teacher memory bank to store features that are more likely to be consistent with the source distribution. Through multimodal prototype alignment and cross-modal relative consistency our method can effectively alleviate domain shift in videos. We evaluate the proposed model on four public video datasets. The results show that our model outperforms existing state-of-the-art methods.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionDomain AdaptationTest-time AdaptationUnsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

Test-Time Adaptation for Nighttime Color-Thermal Semantic Segmentation

2023-07-10 · Yexin Liu, Weiming Zhang, Guoyang Zhao, Jinjing Zhu 외

The ability to scene understanding in adverse visual conditions, e.g., nighttime, has sparked active research for RGB-Thermal (RGB-T) semantic segmentation. However, it is essentially hampered by two critical problems: 1…

Scene UnderstandingSemantic SegmentationTest-time Adaptation

Modality Compensation Network: Cross-Modal Adaptation for Action Recognition

2020-01-31 · Sijie Song, Jiaying Liu, Yanghao Li, Zongming Guo

With the prevalence of RGB-D cameras, multi-modal video data have become more available for human action recognition. One main challenge for this task lies in how to effectively leverage their complementary information. …

Action RecognitionOptical Flow EstimationRepresentation LearningTemporal Action Localization

Retrieve-then-Adapt: Retrieval-Augmented Test-Time Adaptation for Sequential Recommendation

2026-04-07 · Xing Tang, Jingyang Bin, Ziqiang Cui, Xiaokun Zhang 외 arxiv

The sequential recommendation (SR) task aims to predict the next item based on users' historical interaction sequences. Typically trained on historical data, SR models often struggle to adapt to real-time preference shif…

Sequential RecommendationTest-time Adaptation

Unsupervised Domain Adaptation for Cross-Modality Retinal Vessel Segmentation via Disentangling Representation Style Transfer and Collaborative Consistency Learning

2022-01-13 · Linkai Peng, Li Lin, Pujin Cheng, Ziqi Huang 외

Various deep learning models have been developed to segment anatomical structures from medical images, but they typically have poor performance when tested on another target domain with different data distribution. Recen…

Domain AdaptationImage ReconstructionRetinal Vessel SegmentationSegmentation+2

Test-Time Adaptation for Tactile-Vision-Language Models

2026-01-31 · Chuyang Ye, Haoxian Jing, Qinting Jiang, Yixi Lin 외 arxiv

Tactile-vision-language (TVL) models are increasingly deployed in real-world robotic and multimodal perception tasks, where test-time distribution shifts are unavoidable. Existing test-time adaptation (TTA) methods provi…

Test-time Adaptation