Researcher profile
Jenny Zhang
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Quality Diversity through Human Feedback: Towards Open-Ended Diversity-Driven Optimization
2023 · arXiv (Cornell University)
Reinforcement Learning from Human Feedback (RLHF) has shown potential in qualitative tasks where easily defined performance measures are lacking. However, there are drawbacks when RLHF is commonly used to optimize for average human preferences, especially …