ملف الباحث

Xiaozhe Ren

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. EfficientBERT: Progressively Searching Multilayer Perceptron via Warm-up Knowledge Distillation

    2021 · arXiv (Cornell University)

    Pre-trained language models have shown remarkable results on various NLP tasks. Nevertheless, due to their bulky size and slow inference speed, it is hard to deploy them on edge devices. In this paper, we have …

  2. PanGu-$α$: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation

    2021 · arXiv (Cornell University)

    Large-scale Pretrained Language Models (PLMs) have become the new paradigm for Natural Language Processing (NLP). PLMs with hundreds of billions parameters such as GPT-3 have demonstrated strong performances on natural language understanding and generation with …