ملف الباحث
Zhiyang Xu
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
AMELI: Enhancing Multimodal Entity Linking with Fine-Grained Attributes
2023 · arXiv (Cornell University)
We propose attribute-aware multimodal entity linking, where the input consists of a mention described with a text paragraph and images, and the goal is to predict the corresponding target entity from a multimodal knowledge base …
-
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
2025 · arXiv (Cornell University)
Unifying image understanding and generation has gained growing attention in recent research on multimodal models. Although design choices for image understanding have been extensively studied, the optimal model architecture and training recipe for a unified …