Researcher profile
Seungwook Han
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Value Augmented Sampling for Language Model Alignment and Personalization
2024 · ArXiv.org
Aligning Large Language Models (LLMs) to cater to different human preferences, learning new skills, and unlearning harmful behavior is an important problem. Search-based methods, such as Best-of-N or Monte-Carlo Tree Search, are performant, but impractical …