conference-paper
Open access
Refer-iTTS: A System for Referring in Spoken Installments to Objects in Real-World Images
Research footprint
At a glance
- Citations
- 0
- References
- 10
- Comments
- 0
Paper overview
Öz
Current referring expression generation systems mostly deliver their output as one-shot, written expressions. We present on-going work on incremental generation of spoken expressions referring to objects in real-world images. This approach extends upon previous work using the words-as-classifier model for generation. We implement this generator in an incremental dialogue processing framework such that we can exploit an existing interface to incremental text-to-speech synthesis. Our system generates and synthesizes referring expressions while continuously observing non-verbal user reactions.
Record transparency
Publication details
- DOI
- 10.18653/v1/w17-3509
- OpenAlex
- W2755245431
- Document type
- conference-paper
- Language
- EN
- Last metadata update
Comments
Oturum Açın to join the discussion.