Researcher profile

Hyunji Nam

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Efficient RL for optimizing conversation level outcomes with an LLM-based tutor

    2025 · arXiv (Cornell University)

    Large language models (LLMs) built on existing reinforcement learning with human feedback (RLHF) frameworks typically optimize responses based on immediate turn-level human preferences. However, this approach falls short in multi-turn dialogue settings, such as online …