Detection of Glottal Closure Instants From Speech Signals: A Quantitative Review

被引:186
|
作者
Drugman, Thomas [1 ]
Thomas, Mark [2 ]
Gudnason, Jon [3 ]
Naylor, Patrick [2 ]
Dutoit, Thierry [1 ]
机构
[1] Univ Mons, TCTS Lab, B-7000 Mons, Belgium
[2] Univ London Imperial Coll Sci Technol & Med, Dept Elect & Elect Engn, London SW7 2AZ, England
[3] Reykjavik Univ, Sch Sci & Engn, IS-103 Reykjavik, Iceland
关键词
Glottal closure instant (GCI); pitch-synchronous; speech analysis; speech processing; WAVE-FORM; LINEAR PREDICTION; EPOCH EXTRACTION; MODEL;
D O I
10.1109/TASL.2011.2170835
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
The pseudo-periodicity of voiced speech can be exploited in several speech processing applications. This requires however that the precise locations of the glottal closure instants (GCIs) are available. The focus of this paper is the evaluation of automatic methods for the detection of GCIs directly from the speech waveform. Five state-of-the-art GCI detection algorithms are compared using six different databases with contemporaneous electroglottographic recordings as ground truth, and containing many hours of speech by multiple speakers. The five techniques compared are the Hilbert Envelope-based detection (HE), the Zero Frequency Resonator-based method (ZFR), the Dynamic Programming Phase Slope Algorithm (DYPSA), the Speech Event Detection using the Residual Excitation And a Mean-based Signal (SEDREAMS) and the Yet Another GCI Algorithm (YAGA). The efficacy of these methods is first evaluated on clean speech, both in terms of reliabililty and accuracy. Their robustness to additive noise and to reverberation is also assessed. A further contribution of the paper is the evaluation of their performance on a concrete application of speech processing: the causal-anticausal decomposition of speech. It is shown that for clean speech, SEDREAMS and YAGA are the best performing techniques, both in terms of identification rate and accuracy. ZFR and SEDREAMS also show a superior robustness to additive noise and reverberation.
引用
收藏
页码:994 / 1006
页数:13
相关论文
共 50 条
  • [41] DETECTION OF THE GLOTTAL CLOSURE BY JUMPS IN THE STATISTICAL PROPERTIES OF THE SPEECH SIGNAL
    MOULINES, E
    DIFRANCESCO, R
    SPEECH COMMUNICATION, 1990, 9 (5-6) : 401 - 418
  • [42] DETERMINATION OF INSTANT OF GLOTTAL CLOSURE FROM SPEECH WAVE
    STRUBE, HW
    JOURNAL OF THE ACOUSTICAL SOCIETY OF AMERICA, 1974, 56 (05): : 1625 - 1629
  • [43] Extraction of Glottal Closure and Opening Instants using Zero Frequency Filtering
    Deepak, K. T.
    Ramesh, K.
    Prasanna, S. R. M.
    2014 ANNUAL IEEE INDIA CONFERENCE (INDICON), 2014,
  • [44] F0 Variability Measures Based on Glottal Closure Instants
    Chien, Yu-Ren
    Borsky, Michal
    Gudnason, Jon
    INTERSPEECH 2019, 2019, : 1986 - 1989
  • [45] A COMPARISON OF CONVOLUTIONAL NEURAL NETWORKS FOR GLOTTAL CLOSURE INSTANT DETECTION FROM RAW SPEECH
    Matousek, Jindrich
    Tihelka, Daniel
    2021 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP 2021), 2021, : 6938 - 6942
  • [46] SUBBAND ANALYSIS OF LINEAR PREDICTION RESIDUAL FOR THE ESTIMATION OF GLOTTAL CLOSURE INSTANTS
    Vikram, R. L.
    Girish, K. V. Vijay
    Harshavardhan, S.
    Ramakrishnan, A. G.
    Ananthapadmanabhao, T. V.
    2014 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP), 2014,
  • [47] GMAT: Glottal closure instants detection based on the Multiresolution Absolute Teager-Kaiser energy operator
    Wu, Kebin
    Zhang, David
    Lu, Guangming
    DIGITAL SIGNAL PROCESSING, 2017, 69 : 286 - 299
  • [48] Detection of Glottal Opening Instants using Hilbert Envelope
    Ramesh, K.
    Prasanna, S. R. M.
    Govind, D.
    14TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2013), VOLS 1-5, 2013, : 44 - 48
  • [49] Glottal opening instants detection using zero frequency resonator
    Ramesh K.
    Prasanna S.R.M.
    Ramesh, K. (kk.ramesh@iitg.ernet.in), 1600, Springer Science and Business Media, LLC (20): : 127 - 141
  • [50] Midprediction error filtering approach to the detection of glottal closing instants
    Dandapat, S
    Ray, GC
    PROCEEDINGS OF THE 18TH ANNUAL INTERNATIONAL CONFERENCE OF THE IEEE ENGINEERING IN MEDICINE AND BIOLOGY SOCIETY, VOL 18, PTS 1-5, 1997, 18 : 1528 - 1529