Speech recognition using energy, MFCCs and rho parameters to classify syllables in the Spanish language

Sergio Suárez Guerra, José Luis Oropeza Rodriguez, Edgardo Manuel Felipe Riveron, Jesús Figueroa Nazuno

Producción científica: Capítulo del libro/informe/acta de congresoContribución a la conferenciarevisión exhaustiva

Resumen

This paper presents an approach for the automatic speech recognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO and MFCCs parameter increases speech recognition by 5.5% when compared with recognition using STTEF in discontinuous speech and improved more than 2% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 98% for the discontinuous speech and to 81 % for the continuous one.

Idioma originalInglés
Título de la publicación alojadaMICAI 2006
Subtítulo de la publicación alojadaAdvances in Artificial Intelligence - 5th Mexican International Conference on Artificial Intelligence, Proceedings
EditorialSpringer Verlag
Páginas1057-1066
Número de páginas10
ISBN (versión impresa)3540490264, 9783540490265
DOI
EstadoPublicada - 2006
Evento5th Mexican International Conference on Artificial Intelligence, MICAI 2006: Advances in Artificial Intelligence - Apizaco, México
Duración: 13 nov. 200617 nov. 2006

Serie de la publicación

NombreLecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
Volumen4293 LNAI
ISSN (versión impresa)0302-9743
ISSN (versión digital)1611-3349

Conferencia

Conferencia5th Mexican International Conference on Artificial Intelligence, MICAI 2006: Advances in Artificial Intelligence
País/TerritorioMéxico
CiudadApizaco
Período13/11/0617/11/06

Huella

Profundice en los temas de investigación de 'Speech recognition using energy, MFCCs and rho parameters to classify syllables in the Spanish language'. En conjunto forman una huella única.

Citar esto