Speech recognition using energy, MFCCs and Rho parameters to classify syllables in the Spanish language

被引：0

作者：

Suarez Guerra, Sergio ^{[1
]}

Oropeza Rodriguez, Jose Luis ^{[1
]}

Felipe Riveron, Edgardo Manuel ^{[1
]}

Figueroa Nazuno, Jesus ^{[1
]}

机构：

[1] Natl Polytech Inst, Comp Res Ctr, Juan de Dios Batiz S-N, Mexico City 07038, DF, Mexico

来源：

MICAI 2006: ADVANCES IN ARTIFICIAL INTELLIGENCE, PROCEEDINGS | 2006年 / 4293卷

关键词：

D O I：

暂无

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

This paper presents an approach for the automatic speech recognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO and MFCCs parameter increases speech recognition by 5.5% when compared with recognition using STTEF in discontinuous speech and improved more than 2% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 98% for the discontinuous speech and to 81% for the continuous one.

引用

页码：1057 / +

页数：3

共 50 条

[41] Cross-language Transfer Speech Recognition using Deep Learning
Zhao, Yue
Xu, Yan M.
Sun, Mei J.
Xu, Xiao N.
Wang, Hui
Yang, Guo S.
Ji, Qiang
[J]. 11TH IEEE INTERNATIONAL CONFERENCE ON CONTROL AND AUTOMATION (ICCA), 2014, : 1422 - 1426
[42] Identifying Language Disorder in Bilingual Children Using Automatic Speech Recognition
Albudoor, Nahar
Pena, Elizabeth D.
[J]. JOURNAL OF SPEECH LANGUAGE AND HEARING RESEARCH, 2022, 65 (07): : 2648 - 2661
[43] Large vocabulary speech recognition of Slovenian language using morphological models
Maucec, M
Rotovnik, T
Kacic, Z
Horvat, B
[J]. IEEE REGION 8 EUROCON 2003, VOL B, PROCEEDINGS: COMPUTER AS A TOOL, 2003, : 158 - 161
[44] Recognition of target domain Japanese speech using language model replacement
Mori, Daiki
Ohta, Kengo
Nishimura, Ryota
Ogawa, Atsunori
Kitaoka, Norihide
[J]. EURASIP JOURNAL ON AUDIO SPEECH AND MUSIC PROCESSING, 2024, 2024 (01):
[45] Automatic Language Identification Using Speech Rhythm Features for Multi-Lingual Speech Recognition
Kim, Hwamin
Park, Jeong-Sik
[J]. APPLIED SCIENCES-BASEL, 2020, 10 (07):
[46] Teager energy based feature parameters for robust speech recognition in car noise
Jabloun, Firas
Cetin, A.Enis
[J]. ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings, 1999, 1 : 273 - 276
[47] The Teager Energy based feature parameters for robust speech recognition in car noise
Jabloun, F
Çetin, AE
[J]. ICASSP '99: 1999 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, PROCEEDINGS VOLS I-VI, 1999, : 273 - 276
[48] Recognition of Spanish vowels through imagined speech by using spectral analysis and SVM
[J]. 1600, Ubiquitous International (07):
[49] Recognition of vowel segments in Spanish esophageal speech using Hidden Markov Models
Mantilla, Alfredo
Perez-Meana, Hector
Mata, Daniel
Angeles, Carlos
Alvarado, Jorge
Cabrera, Laura
[J]. CIC 2006: 15TH INTERNATIONAL CONFERENCE ON COMPUTING, PROCEEDINGS, 2006, : 115 - +
[50] Speech Enhancement via Smart Larynx of Variable Frequency for Laryngectomee Patient for Tamil Language Syllables Using RADWT Algorithm
Malathi, P.
Suresh, G. R.
Moorthi, M.
Shanker, N. R.
[J]. CIRCUITS SYSTEMS AND SIGNAL PROCESSING, 2019, 38 (09) : 4202 - 4228

← 1 2 3 4 5 →