Detection of Glottal Closure Instants From Speech Signals: A Quantitative Review

被引:186
|
作者
Drugman, Thomas [1 ]
Thomas, Mark [2 ]
Gudnason, Jon [3 ]
Naylor, Patrick [2 ]
Dutoit, Thierry [1 ]
机构
[1] Univ Mons, TCTS Lab, B-7000 Mons, Belgium
[2] Univ London Imperial Coll Sci Technol & Med, Dept Elect & Elect Engn, London SW7 2AZ, England
[3] Reykjavik Univ, Sch Sci & Engn, IS-103 Reykjavik, Iceland
关键词
Glottal closure instant (GCI); pitch-synchronous; speech analysis; speech processing; WAVE-FORM; LINEAR PREDICTION; EPOCH EXTRACTION; MODEL;
D O I
10.1109/TASL.2011.2170835
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
The pseudo-periodicity of voiced speech can be exploited in several speech processing applications. This requires however that the precise locations of the glottal closure instants (GCIs) are available. The focus of this paper is the evaluation of automatic methods for the detection of GCIs directly from the speech waveform. Five state-of-the-art GCI detection algorithms are compared using six different databases with contemporaneous electroglottographic recordings as ground truth, and containing many hours of speech by multiple speakers. The five techniques compared are the Hilbert Envelope-based detection (HE), the Zero Frequency Resonator-based method (ZFR), the Dynamic Programming Phase Slope Algorithm (DYPSA), the Speech Event Detection using the Residual Excitation And a Mean-based Signal (SEDREAMS) and the Yet Another GCI Algorithm (YAGA). The efficacy of these methods is first evaluated on clean speech, both in terms of reliabililty and accuracy. Their robustness to additive noise and to reverberation is also assessed. A further contribution of the paper is the evaluation of their performance on a concrete application of speech processing: the causal-anticausal decomposition of speech. It is shown that for clean speech, SEDREAMS and YAGA are the best performing techniques, both in terms of identification rate and accuracy. ZFR and SEDREAMS also show a superior robustness to additive noise and reverberation.
引用
收藏
页码:994 / 1006
页数:13
相关论文
共 50 条
  • [1] Classification-Based Detection of Glottal Closure Instants from Speech Signals
    Matousek, Jindrich
    Tihelka, Daniel
    18TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2017), VOLS 1-6: SITUATED INTERACTION, 2017, : 3053 - 3057
  • [2] Analysis of algorithms to estimate glottal closure instants from speech signals
    G. Anushiya Rachel
    P. Vijayalakshmi
    T. Nagarajan
    International Journal of Speech Technology, 2020, 23 : 825 - 849
  • [3] Analysis of algorithms to estimate glottal closure instants from speech signals
    Rachel, G. Anushiya
    Vijayalakshmi, P.
    Nagarajan, T.
    INTERNATIONAL JOURNAL OF SPEECH TECHNOLOGY, 2020, 23 (04) : 825 - 849
  • [4] Detection of Glottal Closure Instants from Speech Signals: A Convolutional Neural Network Based Method
    Yang, Shuai
    Wu, Zhiyong
    Shen, Binbin
    Meng, Helen
    19TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2018), VOLS 1-6: SPEECH RESEARCH FOR EMERGING MARKETS IN MULTILINGUAL SOCIETIES, 2018, : 317 - 321
  • [5] COMPARISON OF GLOTTAL CLOSURE INSTANTS DETECTION ALGORITHMS FOR EMOTIONAL SPEECH
    Kadiri, Sudarsana Reddy
    Alku, Paavo
    Yegnanarayana, B.
    2020 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, 2020, : 7379 - 7383
  • [6] Detection of Glottal Closure Instants from Voiced Speech Signals using the Fourier-Bessel Series Expansion
    Mathur, Archit
    Chaudhary, Naveen
    Upadhyay, Abhay
    Pachari, Ram Bilas
    2015 INTERNATIONAL CONFERENCE ON COMMUNICATIONS AND SIGNAL PROCESSING (ICCSP), 2015, : 474 - 478
  • [7] Detection of Glottal Closure Instants from Raw Speech using Convolutional Neural Networks
    Goyal, Mohit
    Srivastava, Varun
    Prathosh, A. P.
    INTERSPEECH 2019, 2019, : 1591 - 1595
  • [8] Glottal Closure and Opening Instant Detection from Speech Signals
    Drugman, Thomas
    Dutoit, Thierry
    INTERSPEECH 2009: 10TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION 2009, VOLS 1-5, 2009, : 2859 - 2862
  • [9] Determination of the instants of glottal closure from speech wave using wavelet transform
    Du, LM
    Hou, ZQ
    ICSP '96 - 1996 3RD INTERNATIONAL CONFERENCE ON SIGNAL PROCESSING, PROCEEDINGS, VOLS I AND II, 1996, : 268 - 271
  • [10] Determination of glottal closure instants from clean and telephone quality speech signals using single frequency filtering
    Kadiri, Sudarsana Reddy
    Yegnanarayana, B.
    COMPUTER SPEECH AND LANGUAGE, 2020, 64