ملف الباحث

Kianté Brantley

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy Optimization

    2022 · arXiv (Cornell University)

    We tackle the problem of aligning pre-trained large language models (LMs) with human preferences. If we view text generation as a sequential decision-making problem, reinforcement learning (RL) appears to be a natural conceptual framework. However, …

  2. Non-Monotonic Sequential Text Generation

    2019 · arXiv (Cornell University)

    Standard sequential generation methods assume a pre-specified generation order, such as text generation methods which generate words from left to right. In this work, we propose a framework for training models of text generation that …