preprint Open access

An Ensemble Method to Produce High-Quality Word Embeddings (2016)

  • arXiv (Cornell University)
  • Cornell University
Research footprint

At a glance

Citations
49
References
28
Comments
0
Paper overview

Abstract

A currently successful approach to computational semantics is to represent words as embeddings in a machine-learned vector space. We present an ensemble method that combines embeddings produced by GloVe (Pennington et al., 2014) and word2vec (Mikolov et al., 2013) with structured knowledge from the semantic networks ConceptNet (Speer and Havasi, 2012) and PPDB (Ganitkevitch et al., 2013), merging their information into a common representation with a large, multilingual vocabulary. The embeddings it produces achieve state-of-the-art performance on many word-similarity evaluations. Its score of $ρ= .596$ on an evaluation of rare words (Luong et al., 2013) is 16% higher than the previous best known system.

Record transparency

Publication details

DOI
10.48550/arxiv.1604.01692
OpenAlex
W2341557172
Document type
preprint
Language
EN
Source
arXiv (Cornell University)
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.