Comparison of glottal closure instants detection algorithms for emotional speech

Sudarsana Kadiri, Paavo Alku, Bayya Yegnanarayana

Tutkimustuotos: Artikkeli kirjassa/konferenssijulkaisussaConference contributionScientificvertaisarvioitu

8 Sitaatiot (Scopus)
136 Lataukset (Pure)

Abstrakti

In production of voiced speech, epochs or glottal closure instants (GCIs) refer to the instants of significant excitation of the vocal tract. Extraction of GCIs is used as a pre-processing stage in many areas of speech technology, such as in prosody modification, speech synthesis and voice source analysis. In the past decades, several GCI detection algorithms have been developed and most of them provide excellent results for speech signals produced using modal (normal) type of phonation. There are, however, no studies comparing multiple state-of-the-art GCI detection methods in emotional speech. In this paper, we compare six GCI detection algorithms using emotional speech and known evaluation metrics. We use the Berlin EMO-DB acted emotional speech database which contains seven emotions and simultaneous electroglottography (EGG) recordings as ground truth. The results show that all six GCI detection algorithms give best performance in processing speech of neutral emotion and that the performance degrade particularly in emotions of high arousal (anger and joy). To improve the performance of GCI detection in emotional speech, the study underlines the importance of local average pitch period estimates.

AlkuperäiskieliEnglanti
Otsikko2020 IEEE International Conference on Acoustics, Speech, and Signal Processing, ICASSP 2020 - Proceedings
KustantajaIEEE
Sivut7379-7383
Sivumäärä5
ISBN (elektroninen)978-1-5090-6631-5
ISBN (painettu)978-1-5090-6631-5
DOI - pysyväislinkit
TilaJulkaistu - toukok. 2020
OKM-julkaisutyyppiA4 Artikkeli konferenssijulkaisussa
TapahtumaIEEE International Conference on Acoustics, Speech, and Signal Processing - Virtual conference, Barcelona, Espanja
Kesto: 4 toukok. 20208 toukok. 2020
Konferenssinumero: 45

Julkaisusarja

NimiProceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing
ISSN (painettu)1520-6149
ISSN (elektroninen)2379-190X

Conference

ConferenceIEEE International Conference on Acoustics, Speech, and Signal Processing
LyhennettäICASSP
Maa/AlueEspanja
KaupunkiBarcelona
Ajanjakso04/05/202008/05/2020
MuuVirtual conference

Sormenjälki

Sukella tutkimusaiheisiin 'Comparison of glottal closure instants detection algorithms for emotional speech'. Ne muodostavat yhdessä ainutlaatuisen sormenjäljen.

Siteeraa tätä