Researcher profile

Longxu Dou

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Efficient Process Reward Model Training via Active Learning

    2025 · arXiv (Cornell University)

    Process Reward Models (PRMs) provide step-level supervision to large language models (LLMs), but scaling up training data annotation remains challenging for both humans and LLMs. To address this limitation, we propose an active learning approach, …