ملف الباحث
Xiaozhe Ren
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
EfficientBERT: Progressively Searching Multilayer Perceptron via Warm-up Knowledge Distillation
2021 · arXiv (Cornell University)
Pre-trained language models have shown remarkable results on various NLP tasks. Nevertheless, due to their bulky size and slow inference speed, it is hard to deploy them on edge devices. In this paper, we have …
-
PanGu-$α$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation
2021 · arXiv (Cornell University)
Large-scale Pretrained Language Models (PLMs) have become the new paradigm for Natural Language Processing (NLP). PLMs with hundreds of billions parameters such as GPT-3 have demonstrated strong performances on natural language understanding and generation with …