ملف الباحث
Lexie Wang
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Filtered Corpus Training (FiCT) Shows that Language Models can Generalize from Indirect Evidence
2024 · arXiv (Cornell University)
This paper introduces Filtered Corpus Training, a method that trains language models (LMs) on corpora with certain linguistic constructions filtered out from the training data, and uses it to measure the ability of LMs to …