ملف الباحث
Armando Solar-Lezama
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Program Synthesis Guided Reinforcement Learning
2021 · arXiv (Cornell University)
A key challenge for reinforcement learning is solving long-horizon planning and control problems. Recent work has proposed leveraging programs to help guide the learning algorithm in these settings. However, these approaches impose a high manual …
-
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
2024 · arXiv (Cornell University)
Large Language Models (LLMs) applied to code-related applications have emerged as a prominent field, attracting significant interest from both academia and industry. However, as new and improved LLMs are developed, existing evaluation benchmarks (e.g., HumanEval, …