Speech recognition using energy, MFCCs and Rho parameters to classify syllables in the Spanish language

被引:0
|
作者
Suarez Guerra, Sergio [1 ]
Oropeza Rodriguez, Jose Luis [1 ]
Felipe Riveron, Edgardo Manuel [1 ]
Figueroa Nazuno, Jesus [1 ]
机构
[1] Natl Polytech Inst, Comp Res Ctr, Juan de Dios Batiz S-N, Mexico City 07038, DF, Mexico
关键词
D O I
暂无
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
This paper presents an approach for the automatic speech recognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO and MFCCs parameter increases speech recognition by 5.5% when compared with recognition using STTEF in discontinuous speech and improved more than 2% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 98% for the discontinuous speech and to 81% for the continuous one.
引用
收藏
页码:1057 / +
页数:3
相关论文
共 50 条
  • [1] Speech recognition using energy parameters to classify syllables in the Spanish language
    Guerra, SS
    Rodríguez, JLO
    Riveron, EMF
    Nazuno, JF
    [J]. PROGRESS IN PATTERN RECOGNITION, IMAGE ANALYSIS AND APPLICATIONS, PROCEEDINGS, 2005, 3773 : 161 - 170
  • [2] Algorithms and Methods for the Automatic Speech Recognition in Spanish Language using Syllables
    Oropeza Rodriguez, Jose Luis
    Suarez Guerra, Sergio
    [J]. COMPUTACION Y SISTEMAS, 2006, 9 (03): : 270 - 286
  • [3] Speech Recognition System Based on OLLO French Corpus by Using MFCCs
    Youcef, Braham Chaouche
    Elemine, Yessaad Mohamed
    Islam, Benmaiza
    Farid, Bouttout
    [J]. RECENT ADVANCES IN ELECTRICAL ENGINEERING AND CONTROL APPLICATIONS, 2017, 411 : 326 - 331
  • [4] Using Syllables as Acoustic Units for Spontaneous Speech Recognition
    Hejtmanek, Jan
    [J]. TEXT, SPEECH AND DIALOGUE, 2010, 6231 : 299 - 305
  • [5] Vowel Recognition from Telephonic Speech Using MFCCs and Gaussian Mixture Models
    Koolagudi, Shashidhar G.
    Thakur, Sujata Negi
    Barthwal, Anurag
    Singh, Manoj Kumar
    Rawat, Ramesh
    Rao, K. Sreenivasa
    [J]. ECO-FRIENDLY COMPUTING AND COMMUNICATION SYSTEMS, 2012, 305 : 170 - +
  • [6] Automatic Disordered Syllables Repetition Recognition in Continuous Speech Using CWT and Correlation
    Codello, Ireneusz
    Kuniszyk-Jozkowiak, Wieslawa
    Smolka, Elzbieta
    Kobus, Adam
    [J]. PROCEEDINGS OF THE 8TH INTERNATIONAL CONFERENCE ON COMPUTER RECOGNITION SYSTEMS CORES 2013, 2013, 226 : 867 - 876
  • [7] Speech emotion recognition using MFCCs extracted from a mobile terminal based on ETSI front end
    Beritelli, Francesco
    Casale, Salvatore
    Russo, Alessandra
    Serrano, Salvatore
    [J]. 2006 8TH INTERNATIONAL CONFERENCE ON SIGNAL PROCESSING, VOLS 1-4, 2006, : 1607 - +
  • [8] Speech Emotion Recognition Using Fourier Parameters
    Wang, Kunxia
    An, Ning
    Li, Bing Nan
    Zhang, Yanyong
    [J]. IEEE TRANSACTIONS ON AFFECTIVE COMPUTING, 2015, 6 (01) : 69 - 75
  • [9] Speech Recognition using HTK Toolkit for Marathi Language
    Chavan, Supriya S.
    Handore, S. M.
    [J]. 2017 IEEE INTERNATIONAL CONFERENCE ON POWER, CONTROL, SIGNALS AND INSTRUMENTATION ENGINEERING (ICPCSI), 2017, : 1591 - 1597
  • [10] Depression Detection in Arabic Using Speech Language Recognition
    Alsharif, Zainab
    Elhag, Salma
    Alfakeh, Sulhi
    [J]. 2022 7TH INTERNATIONAL CONFERENCE ON DATA SCIENCE AND MACHINE LEARNING APPLICATIONS (CDMA 2022), 2022, : 61 - 66