Published January 1, 2014
| Version v1
Conference paper
Open
ANALYSIS OF EFFECT OF SINGLE-CHANNEL SPEECH-MUSIC SEPARATION USING NMF TO AUTOMATIC SPEECH RECOGNITION
- 1. Bogazici Univ, Bilgisayar Muhendisligi, Istanbul, Turkey
- 2. Bogazici Univ, Elekt Elekt Muhendisligi, Istanbul, Turkey
Description
In this study, single-channel speech source separation is carried out to separate the speech from the background music, which degrades the speech recognition performance especially in broadcast news transcription systems. Since the separation is done using single observation of the source signals, the sources have to be previously modeled using training data. Non-negative Matrix Factorization (NMF) methods are used to model the sources. In order to model the source signals, different training data sets, which contain different music and speech data, are created and the effect of the training data sets are analyzed in this study. The performances of the methods are measured not only using separation performance measure but also with speech recognition performance measures.
Files
bib-d6f993d2-a94f-4e25-88ca-adab5b89dab6.txt
Files
(225 Bytes)
| Name | Size | Download all |
|---|---|---|
|
md5:4b0e2eb4fd4cc94a08370c243230b955
|
225 Bytes | Preview Download |