ملف الباحث
Joshua Green
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent
2025 · arXiv (Cornell University)
Deep-Research agents, which integrate large language models (LLMs) with search tools, have shown success in improving the effectiveness of handling complex queries that require iterative search planning and reasoning over search results. Evaluations on current …