ملف الباحث
Dan Iter
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment
2023 · arXiv (Cornell University)
The quality of texts generated by natural language generation (NLG) systems is hard to measure automatically. Conventional reference-based metrics, such as BLEU and ROUGE, have been shown to have relatively low correlation with human judgments, …
-
Automatic Prompt Optimization with “Gradient Descent” and Beam Search
2023
Large Language Models (LLMs) have shown impressive performance as general purpose agents, but their abilities remain highly dependent on prompts which are hand written with onerous trial-and-error effort. We propose a simple and nonparametric solution …