Researcher profile

Jenny Zhang

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Quality Diversity through Human Feedback: Towards Open-Ended Diversity-Driven Optimization

    2023 · arXiv (Cornell University)

    Reinforcement Learning from Human Feedback (RLHF) has shown potential in qualitative tasks where easily defined performance measures are lacking. However, there are drawbacks when RLHF is commonly used to optimize for average human preferences, especially …