article
Open access
Theory of Neural Language Models (Dagstuhl Seminar 25282)
Research footprint
At a glance
- Citations
- 0
- References
- 0
- Comments
- 0
Paper overview
Abstract
This report documents the program and the outcomes of Dagstuhl Seminar 25282 "Theory of Neural Language Models". The seminar aimed to bring researchers together to lay a foundation for continued work on the theory of neural language models, focusing on questions including: How do transformers, RNNs, other NLMs, and their variants, compare with one another in expressivity and trainability? How do the successes and failures of NLMs predicted by theoretical models manifest in practice? What modifications, or what wholly new architectures, are suggested by the theory?
Record transparency
Publication details
- DOI
- 10.4230/dagrep.15.7.22
- OpenAlex
- W7158453664
- Document type
- article
- Language
- EN
- Source
- DROPS (Schloss Dagstuhl – Leibniz Center for Informatics)
- Last metadata update
Comments
Log in to join the discussion.