Researcher profile

Katarina Slama

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Training language models to follow instructions with human feedback

    2022 · arXiv (Cornell University)

    Making language models bigger does not inherently make them better at following a user's intent. For example, large language models can generate outputs that are untruthful, toxic, or simply not helpful to the user. In …