conference-paper

TVT: Transferable Vision Transformer for Unsupervised Domain Adaptation

  • 2023 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)
Research footprint

At a glance

Citations
140
References
86
Comments
0
Paper overview

Abstract

Unsupervised domain adaptation (UDA) aims to transfer the knowledge learnt from a labeled source domain to an unlabeled target domain. Previous work is mainly built upon convolutional neural networks (CNNs) to learn domain-invariant representations. With the recent exponential increase in applying Vision Transformer (ViT) to vision tasks, the capability of ViT in adapting cross-domain knowledge, however, remains unexplored in the literature. To fill this gap, this paper first comprehensively investigates the performance of ViT on a variety of domain adaptation tasks. Surprisingly, ViT demonstrates superior generalization ability, while the performance can be further improved by incorporating adversarial adaptation. Notwithstanding, directly using CNNs-based adaptation strategies fails to take the advantage of ViT’s intrinsic merits (e.g., attention mechanism and sequential image representation) which play an important role in knowledge transfer. To remedy this, we propose an unified framework, namely Transferable Vision Transformer (TVT), to fully exploit the transferability of ViT for domain adaptation. Specifically, we delicately devise a novel and effective unit, which we term Transferability Adaption Module (TAM). By injecting learned trans- ferabilities into attention blocks, TAM compels ViT focus on both transferable and discriminative features. Besides, we leverage discriminative clustering to enhance feature diversity and separation which are undermined during adversarial domain alignment. To verify its versatility, we perform extensive studies of TVT on four benchmarks and the experimental results demonstrate that TVT attains significant improvements compared to existing state-of-the-art UDA methods.

Record transparency

Publication details

DOI
10.1109/wacv56688.2023.00059
OpenAlex
W4319299795
Document type
conference-paper
Language
EN
Source
2023 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.