Improving Speech Transcription For Mandarin-English Translation

M. Tomalin,M. J. F. Gales,X. A. Liu,K. C. Sim,R. Sinha,L. Wang,P. C. Woodland,K. Yu
DOI: https://doi.org/10.1109/ICASSP.2007.367172
2007-01-01
Abstract:This paper describes the development of the CU-HTK Mandarin Speech-To-Text (STT) system and assesses its performance as part of a transcription-translation pipeline which converts broadcast Mandarin audio into English text. Recent improvements to the STT system are described and these give Character Error Rate (CER) gains of 14.3% absolute for a Broadcast Conversation (BC) task and 5.1% absolute for a Broadcast News (BN) task. The output of these STT systems is then post-processed, so that it consists of sentence-like segments, and translated into English text using a Statistical Machine Translation (SMT) system. The performance of the transcription-translation pipeline is evaluated using the Translation Edit Rate (TER) and BLEU metrics. It is shown that improving both the STT system and the post-STT segmentations can lower the TER scores by up to 5.3% absolute and increase the BLEU scores by up to 2.7% absolute.
What problem does this paper attempt to address?