ملف الباحث

Yicheng Zou

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Pre-Trained Policy Discriminators are General Reward Models

    2025

    We offer a novel perspective on reward modeling by formulating it as a policy discriminator, which quantifies the difference between two policies to generate a reward signal, guiding the training policy towards a target policy …

  2. A Lexicon-Based Graph Neural Network for Chinese NER

    2019

    Tao Gui, Yicheng Zou, Qi Zhang, Minlong Peng, Jinlan Fu, Zhongyu Wei, Xuanjing Huang. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language …