ملف الباحث
Yaohong Ding
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Demystify Self-Attention in Vision Transformers from a Semantic Perspective: Analysis and Application
2022 · arXiv (Cornell University)
Self-attention mechanisms, especially multi-head self-attention (MSA), have achieved great success in many fields such as computer vision and natural language processing. However, many existing vision transformer (ViT) works simply inherent transformer designs from NLP to …