Sergey Edunov
8 papers in the PaperMetrix corpus
Papers by this author
-
Facebook FAIR’s WMT19 News Translation Task Submission
2019
This paper describes Facebook FAIR's submission to the WMT19 shared news translation task. We participate in four language directions, English German and English Russian in both directions. Following our submission from last year, our baseline …
-
Understanding Back-Translation at Scale
2018 · arXiv (Cornell University)
An effective method to improve neural machine translation with monolingual data is to augment the parallel training corpus with back-translations of target language sentences. This work broadens the understanding of back-translation and investigates a number …
-
fairseq: A Fast, Extensible Toolkit for Sequence Modeling
2019
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, Michael Auli. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics (Demonstrations). 2019.
-
Multilingual Denoising Pre-training for Neural Machine Translation
2020 · Transactions of the Association for Computational Linguistics
This paper demonstrates that multilingual denoising pre-training produces significant performance gains across a wide variety of machine translation (MT) tasks. We present mBART—a sequence-to-sequence denoising auto-encoder pre-trained on large-scale monolingual corpora in many languages using …
-
Dense Passage Retrieval for Open-Domain Question Answering
2020
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, Wen-tau Yih. Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP). 2020.
-
Beyond English-Centric Multilingual Machine Translation
2020 · arXiv (Cornell University)
Existing work in translation demonstrated the potential of massively multilingual machine translation by training a single model able to translate between any pair of languages. However, much of this work is English-Centric by training only …
-
No Language Left Behind: Scaling Human-Centered Machine Translation
2022 · arXiv (Cornell University)
Driven by the goal of eradicating language barriers on a global scale, machine translation has solidified itself as a key focus of artificial intelligence research today. However, such efforts have coalesced around a small subset …
-
Llama 2: Open Foundation and Fine-Tuned Chat Models
2023 · arXiv (Cornell University)
In this work, we develop and release Llama 2, a collection of pretrained and fine-tuned large language models (LLMs) ranging in scale from 7 billion to 70 billion parameters. Our fine-tuned LLMs, called Llama 2-Chat, …