Zheng-Jun Zha
5 papers in the PaperMetrix corpus
Papers by this author
-
Knowing User Better: Jointly Predicting Click-Through and Playtime for Micro-Video
2019
Most micro-video recommender systems use the click-through to measure user satisfaction. However, the amount of time that users spend on a video, the playtime, measures user engagement on video contents and should be used as …
-
Self-Supervised Visual Representations Learning by Contrastive Mask Prediction
2021 · 2021 IEEE/CVF International Conference on Computer Vision (ICCV)
Advanced self-supervised visual representation learning methods rely on the instance discrimination (ID) pretext task. We point out that the ID task has an implicit semantic consistency (SC) assumption, which may not hold in unconstrained datasets. …
-
ECENet: Explainable and Context-Enhanced Network for Muti-modal Fact verification
2023
Recently, falsified claims incorporating both text and images have been disseminated more effectively than those containing text alone, raising significant concerns for multi-modal fact verification. Existing research makes contributions to multi-modal feature extraction and interaction, …
-
Prototype-Augmented Self-Supervised Generative Network for Generalized Zero-Shot Learning
2024 · IEEE Transactions on Image Processing
Generalized Zero-Shot Learning (GZSL) aims at recognizing images from both seen and unseen classes by constructing correspondences between visual images and semantic embedding. However, existing methods suffer from a strong bias problem, where unseen images …
-
MMAR: Towards Lossless Multi-Modal Auto-Regressive Probabilistic Modeling
2024 · arXiv (Cornell University)
Recent advancements in multi-modal large language models have propelled the development of joint probabilistic models capable of both image understanding and generation. However, we have identified that recent methods suffer from loss of image information …