ملف الباحث

Zheng-Jun Zha

5 أوراق في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Knowing User Better: Jointly Predicting Click-Through and Playtime for Micro-Video

    2019

    Most micro-video recommender systems use the click-through to measure user satisfaction. However, the amount of time that users spend on a video, the playtime, measures user engagement on video contents and should be used as …

  2. Self-Supervised Visual Representations Learning by Contrastive Mask Prediction

    2021 · 2021 IEEE/CVF International Conference on Computer Vision (ICCV)

    Advanced self-supervised visual representation learning methods rely on the instance discrimination (ID) pretext task. We point out that the ID task has an implicit semantic consistency (SC) assumption, which may not hold in unconstrained datasets. …

  3. ECENet: Explainable and Context-Enhanced Network for Muti-modal Fact verification

    2023

    Recently, falsified claims incorporating both text and images have been disseminated more effectively than those containing text alone, raising significant concerns for multi-modal fact verification. Existing research makes contributions to multi-modal feature extraction and interaction, …

  4. Prototype-Augmented Self-Supervised Generative Network for Generalized Zero-Shot Learning

    2024 · IEEE Transactions on Image Processing

    Generalized Zero-Shot Learning (GZSL) aims at recognizing images from both seen and unseen classes by constructing correspondences between visual images and semantic embedding. However, existing methods suffer from a strong bias problem, where unseen images …

  5. MMAR: Towards Lossless Multi-Modal Auto-Regressive Probabilistic Modeling

    2024 · arXiv (Cornell University)

    Recent advancements in multi-modal large language models have propelled the development of joint probabilistic models capable of both image understanding and generation. However, we have identified that recent methods suffer from loss of image information …