Researcher profile
Q Zhang
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
When Human Preferences Flip: An Instance-Dependent Robust Loss for RLHF
2025 · arXiv (Cornell University)
Quality of datasets plays an important role in large language model (LLM) alignment. In collecting human feedback, however, preference flipping is ubiquitous and causes corruption in data annotation; the issue necessitates the alignment algorithms with …