ملف الباحث
Rui Zhong
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Simplified Multiplayer Battle Game-inspired Optimizer with Diverse Search Strategies
2024
We propose two modifications to the standard multiplayer battle game-inspired optimizer (MBGO) to simplify its search framework and enhance its performance. Specifically, the first modification changes the original serial two-stage search to a parallel approach, …
-
Navigate the Unknown: Enhancing LLM Reasoning with Intrinsic Motivation Guided Exploration
2025 · arXiv (Cornell University)
Reinforcement Learning (RL) has become a key approach for enhancing the reasoning capabilities of large language models. However, prevalent RL approaches like proximal policy optimization and group relative policy optimization suffer from sparse, outcome-based rewards …