ملف الباحث

Guan, XinYi

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Web-Bench: A LLM Code Benchmark Based on Web Standards and Frameworks

    2025 · arXiv (Cornell University)

    The application of large language models (LLMs) in the field of coding is evolving rapidly: from code assistants, to autonomous coding agents, and then to generating complete projects through natural language. Early LLM code benchmarks …