Researcher profile

Feng-Lin Li

2 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Don't Forget Your Reward Values: Language Model Alignment via Value-based Calibration

    2024 · arXiv (Cornell University)

    While Reinforcement Learning from Human Feedback (RLHF) significantly enhances the generation quality of Large Language Models (LLMs), recent studies have raised concerns regarding the complexity and instability associated with the Proximal Policy Optimization (PPO) algorithm, …

  2. AliMe Chat: A Sequence to Sequence and Rerank based Chatbot Engine

    2017

    Minghui Qiu, Feng-Lin Li, Siyu Wang, Xing Gao, Yan Chen, Weipeng Zhao, Haiqing Chen, Jun Huang, Wei Chu. Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers). 2017.