Researcher profile

Yinmin Zhong

2 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Optimizing RLHF Training for Large Language Models with Stage Fusion

    2024 · arXiv (Cornell University)

    We present RLHFuse, an efficient training system with stage fusion for Reinforcement Learning from Human Feedback (RLHF). Due to the intrinsic nature of RLHF training, i.e., the data skewness in the generation stage and the …

  2. LoongServe: Efficiently Serving Long-Context Large Language Models with Elastic Sequence Parallelism

    2024

    The context window of large language models (LLMs) is rapidly increasing, leading to a huge variance in resource usage between different requests as well as between different phases of the same request. Restricted by static …