Researcher profile
Hyunji Nam
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Efficient RL for optimizing conversation level outcomes with an LLM-based tutor
2025 · arXiv (Cornell University)
Large language models (LLMs) built on existing reinforcement learning with human feedback (RLHF) frameworks typically optimize responses based on immediate turn-level human preferences. However, this approach falls short in multi-turn dialogue settings, such as online …