Speech recognition using energy parameters to classify syllables in the Spanish language

Sergio Suárez Guerra; José Luis Oropeza Rodríguez; Edgardo M.Felipe Riveron; Jesús Figueroa Nazuno

doi:10.1007/11578079_18

Speech recognition using energy parameters to classify syllables in the Spanish language

Sergio Suárez Guerra, José Luis Oropeza Rodríguez, Edgardo M.Felipe Riveron, Jesús Figueroa Nazuno

Centro de Investigación en Computación (CIC)

Producción científica: Capítulo del libro/informe/acta de congreso › Contribución a la conferencia › revisión exhaustiva

1 Cita (Scopus)

Resumen

This paper presents an approach for the automatic speech recognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO parameter increases speech recognition by 5% when compared with recognition using STTEF in discontinuous speech and improved more than 1.5% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 97.5% for the discontinuous speech and to 80.5% for the continuous one.

Idioma original	Inglés
Título de la publicación alojada	Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings
Páginas	161-170
Número de páginas	10
DOI	https://doi.org/10.1007/11578079_18
Estado	Publicada - 2005
Evento	10th Iberoamerican Congress on Pattern Recognition, CIARP 2005 - Havana, Cuba Duración: 15 nov. 2005 → 18 nov. 2005

Serie de la publicación

Nombre	Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
Volumen	3773 LNCS
ISSN (versión impresa)	0302-9743
ISSN (versión digital)	1611-3349

Conferencia

Conferencia	10th Iberoamerican Congress on Pattern Recognition, CIARP 2005
País/Territorio	Cuba
Ciudad	Havana
Período	15/11/05 → 18/11/05

Acceder al documento

10.1007/11578079_18

Otros archivos y enlaces

Enlace a la publicación en Scopus

Citar esto

Guerra, S. S., Rodríguez, J. L. O., Riveron, E. M. F., & Nazuno, J. F. (2005). Speech recognition using energy parameters to classify syllables in the Spanish language. En Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings (pp. 161-170). (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics); Vol. 3773 LNCS). https://doi.org/10.1007/11578079_18

Guerra, Sergio Suárez ; Rodríguez, José Luis Oropeza ; Riveron, Edgardo M.Felipe et al. / Speech recognition using energy parameters to classify syllables in the Spanish language. Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings. 2005. pp. 161-170 (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)).

@inproceedings{d6055077c6cb4542a3755a1c10f257e8,

title = "Speech recognition using energy parameters to classify syllables in the Spanish language",

abstract = "This paper presents an approach for the automatic speech recognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO parameter increases speech recognition by 5% when compared with recognition using STTEF in discontinuous speech and improved more than 1.5% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 97.5% for the discontinuous speech and to 80.5% for the continuous one.",

author = "Guerra, {Sergio Su{\'a}rez} and Rodr{\'i}guez, {Jos{\'e} Luis Oropeza} and Riveron, {Edgardo M.Felipe} and Nazuno, {Jes{\'u}s Figueroa}",

year = "2005",

doi = "10.1007/11578079_18",

language = "Ingl{\'e}s",

isbn = "3540298509",

series = "Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)",

pages = "161--170",

booktitle = "Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings",

note = "10th Iberoamerican Congress on Pattern Recognition, CIARP 2005 ; Conference date: 15-11-2005 Through 18-11-2005",

}

Guerra, SS, Rodríguez, JLO , Riveron, EMF & Nazuno, JF 2005, Speech recognition using energy parameters to classify syllables in the Spanish language. En Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings. Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics), vol. 3773 LNCS, pp. 161-170, 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Havana, Cuba, 15/11/05. https://doi.org/10.1007/11578079_18

Speech recognition using energy parameters to classify syllables in the Spanish language. / Guerra, Sergio Suárez; Rodríguez, José Luis Oropeza ; Riveron, Edgardo M.Felipe et al.
Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings. 2005. p. 161-170 (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics); Vol. 3773 LNCS).

Producción científica: Capítulo del libro/informe/acta de congreso › Contribución a la conferencia › revisión exhaustiva

TY - GEN

T1 - Speech recognition using energy parameters to classify syllables in the Spanish language

AU - Guerra, Sergio Suárez

AU - Rodríguez, José Luis Oropeza

AU - Riveron, Edgardo M.Felipe

AU - Nazuno, Jesús Figueroa

PY - 2005

Y1 - 2005

N2 - This paper presents an approach for the automatic speech recognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO parameter increases speech recognition by 5% when compared with recognition using STTEF in discontinuous speech and improved more than 1.5% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 97.5% for the discontinuous speech and to 80.5% for the continuous one.

AB - This paper presents an approach for the automatic speech recognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO parameter increases speech recognition by 5% when compared with recognition using STTEF in discontinuous speech and improved more than 1.5% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 97.5% for the discontinuous speech and to 80.5% for the continuous one.

UR - http://www.scopus.com/inward/record.url?scp=33745367125&partnerID=8YFLogxK

U2 - 10.1007/11578079_18

DO - 10.1007/11578079_18

M3 - Contribución a la conferencia

SN - 3540298509

SN - 9783540298502

T3 - Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)

SP - 161

EP - 170

BT - Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings

T2 - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005

Y2 - 15 November 2005 through 18 November 2005

ER -

Guerra SS, Rodríguez JLO , Riveron EMF, Nazuno JF. Speech recognition using energy parameters to classify syllables in the Spanish language. En Progress in Pattern Recognition, Image Analysis and Applications - 10th Iberoamerican Congress on Pattern Recognition, CIARP 2005, Proceedings. 2005. p. 161-170. (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)). doi: 10.1007/11578079_18

Speech recognition using energy parameters to classify syllables in the Spanish language

Resumen

Serie de la publicación

Conferencia

Acceder al documento

Otros archivos y enlaces

Huella

Citar esto