Ryan Lowe
4 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
A Survey of Available Corpora for Building Data-Driven Dialogue Systems
2015 · arXiv (Cornell University)
During the past decade, several areas of speech and language understanding have witnessed substantial breakthroughs from the use of data-driven models. In the area of dialogue systems, the trend is less obvious, and most practical …
-
How NOT To Evaluate Your Dialogue System: An Empirical Study of Unsupervised Evaluation Metrics for Dialogue Response Generation
2016 · arXiv (Cornell University)
We investigate evaluation metrics for dialogue response generation systems where supervised labels, such as task completion, are not available. Recent works in response generation have adopted metrics from machine translation to compare a model's generated …
-
An Actor-Critic Algorithm for Sequence Prediction
2016 · arXiv (Cornell University)
We present an approach to training neural networks to generate sequences using actor-critic methods from reinforcement learning (RL). Current log-likelihood training methods are limited by the discrepancy between their training and testing modes, as models …
-
Training language models to follow instructions with human feedback
2022 · arXiv (Cornell University)
Making language models bigger does not inherently make them better at following a user's intent. For example, large language models can generate outputs that are untruthful, toxic, or simply not helpful to the user. In …