Douglas Eck
3 أوراق في مجموعة PaperMetrix
أوراق هذا المؤلف
-
Online and Linear-Time Attention by Enforcing Monotonic Alignments
2017 · arXiv (Cornell University)
Recurrent neural network models with an attention mechanism have proven to be extremely effective on a wide variety of sequence-to-sequence problems. However, the fact that soft attention mechanisms perform a pass over the entire input …
-
Deduplicating Training Data Makes Language Models Better
2022 · Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Katherine Lee, Daphne Ippolito, Andrew Nystrom, Chiyuan Zhang, Douglas Eck, Chris Callison-Burch, Nicholas Carlini. Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2022.
-
PaLM: Scaling Language Modeling with Pathways
2022 · arXiv (Cornell University)
Large language models have been shown to achieve remarkable performance across a variety of natural language tasks using few-shot learning, which drastically reduces the number of task-specific training examples needed to adapt the model to …