Marie-Anne Lachaux
3 papers in the PaperMetrix corpus
Papers by this author
-
Poly-encoders: Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence Scoring
2020 · International Conference on Learning Representations
The use of deep pre-trained transformers has led to remarkable progress in a number of applications (Devlin et al., 2018). For tasks that make pairwise comparisons between sequences, matching a given input with a corresponding …
-
LLaMA: Open and Efficient Foundation Language Models
2023 · arXiv (Cornell University)
We introduce LLaMA, a collection of foundation language models ranging from 7B to 65B parameters. We train our models on trillions of tokens, and show that it is possible to train state-of-the-art models using publicly …
-
Llama 2: Open Foundation and Fine-Tuned Chat Models
2023 · arXiv (Cornell University)
In this work, we develop and release Llama 2, a collection of pretrained and fine-tuned large language models (LLMs) ranging in scale from 7 billion to 70 billion parameters. Our fine-tuned LLMs, called Llama 2-Chat, …