Yidong Wang
4 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
SoftMatch: Addressing the Quantity-Quality Trade-off in Semi-supervised Learning
2023 · arXiv (Cornell University)
The critical challenge of Semi-Supervised Learning (SSL) is how to effectively leverage the limited labeled data and massive unlabeled data to improve the model's generalization performance. In this paper, we first revisit the popular pseudo-labeling …
-
Evaluating Open-QA Evaluation
2023 · arXiv (Cornell University)
This study focuses on the evaluation of the Open Question Answering (Open-QA) task, which can directly estimate the factuality of large language models (LLMs). Current automatic evaluation methods have shown limitations, indicating that human evaluation …
-
RAGLAB: A Modular and Research-Oriented Unified Framework for Retrieval-Augmented Generation
2024 · arXiv (Cornell University)
Large Language Models (LLMs) demonstrate human-level capabilities in dialogue, reasoning, and knowledge retention. However, even the most advanced LLMs face challenges such as hallucinations and real-time updating of their knowledge. Current research addresses this bottleneck …
-
A Survey on Evaluation of Large Language Models
2024 · ACM Transactions on Intelligent Systems and Technology
Large language models (LLMs) are gaining increasing popularity in both academia and industry, owing to their unprecedented performance in various applications. As LLMs continue to play a vital role in both research and daily use, …