Sam McCandlish
3 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
Discovering Language Model Behaviors with Model-Written Evaluations
2022 · arXiv (Cornell University)
As language models (LMs) scale, they develop many novel behaviors, good and bad, exacerbating the need to evaluate how they behave. Prior work creates evaluations with crowdwork (which is time-consuming and expensive) or existing data …
-
Scaling Laws for Neural Language Models
2020 · arXiv (Cornell University)
This paper develops a transport-validity theory for agentic AI interventions that are first screened on small systems and later considered for frontier-scale deployment. Rather than predicting absolute frontier performance, it asks when a comparative gain …
-
Language Models are Few-Shot Learners
2020 · arXiv (Cornell University)
Recent work has demonstrated substantial gains on many NLP tasks and benchmarks by pre-training on a large corpus of text followed by fine-tuning on a specific task. While typically task-agnostic in architecture, this method still …