ملف الباحث

Eric Noland

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Training Compute-Optimal Large Language Models

    2022 · arXiv (Cornell University)

    We investigate the optimal model size and number of tokens for training a transformer language model under a given compute budget. We find that current large language models are significantly undertrained, a consequence of the …