conference-paper Open access

Deconvolutional Latent-Variable Model for Text Sequence Matching

  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Association for the Advancement of Artificial Intelligence
Research footprint

At a glance

Citations
59
References
69
Comments
0
Paper overview

Öz

A latent-variable model is introduced for text matching, inferring sentence representations by jointly optimizing generative and discriminative objectives. To alleviate typical optimization challenges in latent-variable models for text, we employ deconvolutional networks as the sequence decoder (generator), providing learned latent codes with more semantic information and better generalization. Our model, trained in an unsupervised manner, yields stronger empirical predictive performance than a decoder based on Long Short-Term Memory (LSTM), with less parameters and considerably faster training. Further, we apply it to text sequence-matching problems. The proposed model significantly outperforms several strong sentence-encoding baselines, especially in the semi-supervised setting.

Record transparency

Publication details

DOI
10.1609/aaai.v32i1.11991
OpenAlex
W2756946152
Document type
conference-paper
Language
EN
Source
Proceedings of the AAAI Conference on Artificial Intelligence
Last metadata update
Community

Comments

Oturum Açın to join the discussion.

  1. No comments yet. Start the discussion.