Study of Formant Modification for Children ASR

Hemant Kathania, Sudarsana Kadiri, Paavo Alku, Mikko Kurimo

Tutkimustuotos: Artikkeli kirjassa/konferenssijulkaisussaConference article in proceedingsScientificvertaisarvioitu

42 Sitaatiot (Scopus)
274 Lataukset (Pure)

Abstrakti

The performance of automatic speech recognition systems for children’s speech is known to suffer from the large variation and mismatch in the acoustic and linguistic attributes between children’s and adults’ speech. One of the various identified sources of mismatch is the difference in formant frequencies between adults and children. In this paper, we propose a formant modification method to mitigate differences between adults’ and children’s speech and to improve the performance of ASR for children. The explored technique gives a relative 27% improvement in system performance compared to a hybrid DNN-HMM baseline. We also compare the system performance with related speaker adaptation methods like vocal tract length normalization (VTLN) and speaking rate adapta-
tion (SRA) and find that the proposed method gives improvements over them, as well. Combining the proposed method with VTLN and SRA results in a further reduction of WER. We also found that the proposed method performs well even
for noisy speech.
AlkuperäiskieliEnglanti
Otsikko2020 IEEE International Conference on Acoustics, Speech, and Signal Processing, ICASSP 2020 - Proceedings
KustantajaIEEE
Sivut7429-7433
Sivumäärä5
ISBN (elektroninen)978-1-5090-6631-5
ISBN (painettu)978-1-5090-6632-2
DOI - pysyväislinkit
TilaJulkaistu - toukok. 2020
OKM-julkaisutyyppiA4 Artikkeli konferenssijulkaisussa
TapahtumaIEEE International Conference on Acoustics, Speech, and Signal Processing - Virtual conference, Barcelona, Espanja
Kesto: 4 toukok. 20208 toukok. 2020
Konferenssinumero: 45

Julkaisusarja

NimiProceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing
ISSN (painettu)1520-6149
ISSN (elektroninen)2379-190X

Conference

ConferenceIEEE International Conference on Acoustics, Speech, and Signal Processing
LyhennettäICASSP
Maa/AlueEspanja
KaupunkiBarcelona
Ajanjakso04/05/202008/05/2020
MuuVirtual conference

Sormenjälki

Sukella tutkimusaiheisiin 'Study of Formant Modification for Children ASR'. Ne muodostavat yhdessä ainutlaatuisen sormenjäljen.

Siteeraa tätä