Published January 1, 2019 | Version v1
Conference paper Open

Improving Low Resource Turkish Speech Recognition with Data Augmentation and TTS

  • 1. TUBITAK BILGEM, Speech & Language Technol Lab, Kocaeli, Turkey
  • 2. Istanbul Tech Univ, Visual Intelligence Lab, Istanbul, Turkey

Description

One of the major problems faced by speech recognition researchers is the lack of data. In this paper, our objective is to compare alternative solutions to lack of data. Some experiments are conducted with very limited training data to see the effects of data augmentation and speech synthesis on speech recognition. Speed and volume perturbations are applied in this study. Besides data augmentation, synthetic speech is generated by using two different speech synthesis methods. In first speech synthesis approach, Google Translate Text to Speech (gTTS) is used as speech synthesizer. In second speech synthesis approach, an end-to-end Turkish TTS system is trained by us. Finally, we examined the effects of all these alternative methods on speech recognition for low resource languages.

Files

bib-9c628267-47e4-4f04-ad4d-026edab3a13a.txt

Files (189 Bytes)

Name Size Download all
md5:4200d6729a96432372b5d6b177670178
189 Bytes Preview Download