Researcher profile

Seungwook Han

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Value Augmented Sampling for Language Model Alignment and Personalization

    2024 · ArXiv.org

    Aligning Large Language Models (LLMs) to cater to different human preferences, learning new skills, and unlearning harmful behavior is an important problem. Search-based methods, such as Best-of-N or Monte-Carlo Tree Search, are performant, but impractical …