Publications
Enriching machine-mediated speech-to-speech translation using contextual information
Abstract
Conventional approaches to speech-to-speech (S2S) translation typically ignore key contextual information such as prosody, emphasis, discourse state in the translation process. Capturing and exploiting such contextual information is especially important in machine-mediated S2S translation as it can serve as a complementary knowledge source that can potentially aid the end users in improved understanding and disambiguation. In this work, we present a general framework for integrating rich contextual information in S2S translation. We present novel methodologies for integrating source side context in the form of dialog act (DA) tags, and target side context using prosodic word prominence. We demonstrate the integration of the DA tags in two different statistical translation frameworks, phrase-based translation and a bag-of-words lexical choice model. In addition to producing interpretable DA annotated target …
- Date
- 2013
- Authors
- Vivek Kumar Rangarajan Sridhar, Srinivas Bangalore, Shrikanth Narayanan
- Journal
- Computer Speech & Language
- Volume
- 27
- Issue
- 2
- Pages
- 492-508
- Publisher
- Academic Press