conference-paper Open access

Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models

  • Findings of the Association for Computational Linguistics: ACL 2022
Research footprint

At a glance

Citations
282
References
42
Comments
0
Paper overview

Abstract

We provide the first exploration of sentence embeddings from text-to-text transformers (T5) including the effects of scaling up sentence encoders to 11B parameters. Sentence embeddings are broadly useful for language processing tasks. While T5 achieves impressive performance on language tasks, it is unclear how to produce sentence embeddings from encoder-decoder models. We investigate three methods to construct Sentence-T5 (ST5) models: two utilize only the T5 encoder and one using the full T5 encoderdecoder. We establish a new sentence representation transfer benchmark, SentGLUE, which extends the SentEval toolkit to nine tasks from the GLUE benchmark Our encoder-only models outperform the previous best models on both SentEval and SentGLUE transfer tasks, including semantic textual similarity (STS). Scaling up ST5 from millions to billions of parameters shown to consistently improve performance. Finally, our encoderdecoder method achieves a new state-of-theart on STS when using sentence embeddings. 1

Record transparency

Publication details

DOI
10.18653/v1/2022.findings-acl.146
OpenAlex
W3194782062
Document type
conference-paper
Language
EN
Source
Findings of the Association for Computational Linguistics: ACL 2022
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.