ملف الباحث
Yiping Lu
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
A Simple "Try Again" Can Elicit Multi-Turn LLM Reasoning
2025 · arXiv (Cornell University)
Multi-turn problem solving is critical yet challenging for Large Reasoning Models (LRMs) to reflect on their reasoning and revise from feedback. Existing Reinforcement Learning (RL) methods train large reasoning models on a single-turn paradigm with …