ملف الباحث
Hugh Zhang
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Chain-of-Thought Reasoning is a Policy Improvement Operator
2023 · arXiv (Cornell University)
Large language models have astounded the world with fascinating new capabilities. However, they currently lack the ability to teach themselves new skills, relying instead on large amounts of human-generated training data. We introduce SECToR (Self-Education …
-
Unifying Human and Statistical Evaluation for Natural Language Generation
2019
Tatsunori B. Hashimoto, Hugh Zhang, Percy Liang. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). 2019.