conference-paper Open access

Improved Transition-based Parsing by Modeling Characters instead of Words with LSTMs

  • RECERCAT (Consorci de Serveis Universitaris de Catalunya)
  • Consorci de Serveis Universitaris de Catalunya
Research footprint

At a glance

Citations
245
References
47
Comments
0
Paper overview

Abstract

We present extensions to a continuousstate dependency parsing method that makes it applicable to morphologically rich languages. Starting with a highperformance transition-based parser that uses long short-term memory (LSTM) recurrent neural networks to learn representations of the parser state, we replace lookup-based word representations with representations constructed from the orthographic representations of the words, also using LSTMs. This allows statistical sharing across word forms that are similar on the surface. Experiments for morphologically rich languages show that the parsing model benefits from incorporating the character-based encodings of words.

Record transparency

Publication details

DOI
10.18653/v1/d15-1041
OpenAlex
W1860935423
Document type
conference-paper
Language
EN
Source
RECERCAT (Consorci de Serveis Universitaris de Catalunya)
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.