Amanda Askell
3 papers in the PaperMetrix corpus
Papers by this author
-
Discovering Language Model Behaviors with Model-Written Evaluations
2022 · arXiv (Cornell University)
As language models (LMs) scale, they develop many novel behaviors, good and bad, exacerbating the need to evaluate how they behave. Prior work creates evaluations with crowdwork (which is time-consuming and expensive) or existing data …
-
Language Models are Few-Shot Learners
2020 · arXiv (Cornell University)
Recent work has demonstrated substantial gains on many NLP tasks and benchmarks by pre-training on a large corpus of text followed by fine-tuning on a specific task. While typically task-agnostic in architecture, this method still …
-
Training language models to follow instructions with human feedback
2022 · arXiv (Cornell University)
Making language models bigger does not inherently make them better at following a user's intent. For example, large language models can generate outputs that are untruthful, toxic, or simply not helpful to the user. In …