ملف الباحث
Xiangyu Peng
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Reducing Non-Normative Text Generation from Language Models
2020 · arXiv (Cornell University)
Large-scale, transformer-based language models such as GPT-2 are pretrained on diverse corpora scraped from the internet. Consequently, they are prone to generating non-normative text (i.e. in violation of social norms). We introduce a technique for …
-
Reliable Label Correction is a Good Booster When Learning with Extremely Noisy Labels
2022 · arXiv (Cornell University)
Learning with noisy labels has aroused much research interest since data annotations, especially for large-scale datasets, may be inevitably imperfect. Recent approaches resort to a semi-supervised learning problem by dividing training samples into clean and …