preprint
Open access
Bridging the data gap between children and large language models
Research footprint
At a glance
- Citations
- 13
- References
- 11
- Comments
- 0
Paper overview
Abstract
Large language models show intriguing emergent behaviors, yet they receive around 4-5 orders of magnitude more language data than human children. What accounts for this vast difference in sample efficiency? Candidate explanations include children’s pre-existing conceptual structures, their use of multimodal grounding, and the interactive, social nature of their input.
Record transparency
Publication details
- DOI
- 10.31234/osf.io/qzbgx
- OpenAlex
- W4382404227
- Document type
- preprint
- Language
- EN
- Last metadata update
Comments
Log in to join the discussion.