ملف الباحث
Yicheng Zou
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Pre-Trained Policy Discriminators are General Reward Models
2025
We offer a novel perspective on reward modeling by formulating it as a policy discriminator, which quantifies the difference between two policies to generate a reward signal, guiding the training policy towards a target policy …
-
A Lexicon-Based Graph Neural Network for Chinese NER
2019
Tao Gui, Yicheng Zou, Qi Zhang, Minlong Peng, Jinlan Fu, Zhongyu Wei, Xuanjing Huang. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language …