conference-paper Open access

An Unsupervised Probability Model for Speech-to-Translation Alignment of Low-Resource Languages

Research footprint

At a glance

Citations
26
References
17
Comments
0
Paper overview

Abstract

For many low-resource languages, spoken language resources are more likely to be annotated with translations than with transcriptions. Translated speech data is potentially valuable for documenting endangered languages or for training speech translation systems. A first step towards making use of such data would be to automatically align spoken words with their translations. We present a model that combines Dyer et al.'s reparameterization of IBM Model 2 (fast-align) and k-means clustering using Dynamic Time Warping as a distance metric. The two components are trained jointly using expectation-maximization. In an extremely low-resource scenario, our model performs significantly better than both a neural model and a strong baseline.

Record transparency

Publication details

DOI
10.18653/v1/d16-1133
OpenAlex
W2964102148
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.