ملف الباحث

Long Ouyang

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Self-critiquing models for assisting human evaluators

    2022 · arXiv (Cornell University)

    We fine-tune large language models to write natural language critiques (natural language critical comments) using behavioral cloning. On a topic-based summarization task, critiques written by our models help humans find flaws in summaries that they …

  2. Training language models to follow instructions with human feedback

    2022 · arXiv (Cornell University)

    Making language models bigger does not inherently make them better at following a user's intent. For example, large language models can generate outputs that are untruthful, toxic, or simply not helpful to the user. In …