ملف الباحث
Yang, Yan
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Towards Alleviating Text-to-Image Retrieval Hallucination for CLIP in Zero-shot Learning
2024 · arXiv (Cornell University)
Pretrained cross-modal models, for instance, the most representative CLIP, have recently led to a boom in using pre-trained models for cross-modal zero-shot tasks, considering the generalization properties. However, we analytically discover that CLIP suffers from …