Speech Recognition Using Energy Parameters to Classify Syllables in the Spanish Language

Guerra, Sergio Suárez; Rodríguez, José Luis Oropeza; Riveron, Edgardo M. Felipe; Nazuno, Jesús Figueroa

doi:10.1007/11578079_18

Sergio Suárez Guerra¹⁸,
José Luis Oropeza Rodríguez¹⁸,
Edgardo M. Felipe Riveron¹⁸ &
…
Jesús Figueroa Nazuno¹⁸

Part of the book series: Lecture Notes in Computer Science ((LNIP,volume 3773))

Included in the following conference series:

Iberoamerican Congress on Pattern Recognition

1159 Accesses

Abstract

This paper presents an approach for the automatic speech re-cognition using syllabic units. Its segmentation is based on using the Short-Term Total Energy Function (STTEF) and the Energy Function of the High Frequency (ERO parameter) higher than 3,5 KHz of the speech signal. Training for the classification of the syllables is based on ten related Spanish language rules for syllable splitting. Recognition is based on a Continuous Density Hidden Markov Models and the bigram model language. The approach was tested using two voice corpus of natural speech, one constructed for researching in our laboratory (experimental) and the other one, the corpus Latino40 commonly used in speech researches. The use of ERO parameter increases speech recognition by 5% when compared with recognition using STTEF in discontinuous speech and improved more than 1.5% in continuous speech with three states. When the number of states is incremented to five, the recognition rate is improved proportionally to 97.5% for the discontinuous speech and to 80.5% for the continuous one.

Download to read the full chapter text

Chapter PDF

Automatic Syllabification and Syllable Timing of Automatically Recognized Speech – for Czech

Automatic Syllable Repetition Detection in Continuous Speech Based on Linear Prediction Coefficients

A study on conventional and syllable-based approaches for automatic speech recognition in Malayalam

Article 20 December 2022

Keywords

These keywords were added by machine and not by the authors. This process is experimental and the keywords may be updated as the learning algorithm improves.

References

Meneido, H., Neto, J.: Combination of Acoustic Models in Continuous Speech Recognition Hybrid Systems, INESC, Rua Alves Redol, 9, 1000- 029 Lisbon, Portugal (2000)
Google Scholar
Meneido, H., Joâo, P., Neto, J., Luis, B., Almeida, L.: INESC-IST. Syllable Onset Detection Applied to the Portuguese Language. In: 6th European Conference on Speech Communication and Technology (EUROSPEECH 1999), Budapest, Hungary, September 5-9 (1999)
Google Scholar
Suárez, S., Oropeza, J.L., Suso, K., del Villar, M.: Pruebas y validación de un sistema de reconocimiento del habla basado en sílabas con un vocabulario pequeño. In: Congreso Internacional de Computación CIC 2003, México, D.F (2003)
Google Scholar
Wu, S.-L., Shire, M.L., Greenberg, S., Morgan, N.: Integrating Syllable Boundary Information into Speech Recognition. In: Proc. ICASSP (1998)
Google Scholar
Rabiner, L., Juang, B.-H.: Fundamentals of Speech Recognition. Prentice Hall, Englewood Cliffs
Google Scholar
Serridge, B.: Análisis del Español Mexicano, para la construcción de un sistema de reconocimiento de dicho lenguaje. In: Grupo TLATOA, UDLA, Puebla, México (1993)
Google Scholar
Fujimura, O.: UCI Working Papers in Linguistics. In: Proceedings of the South Western Optimality Theory Workshop (SWOT II), Syllable Structure Constraints, a C/D Model Perspective, vol. 2 (1996)
Google Scholar
Wu, S.: Incorporating information from syllable-length time scales into automatic speech recognition. PhD Thesis, Berkeley University, California (1998)
Google Scholar
Bilmes, J.A.: A Gentle Tutorial of the EM Algorithm and its Application to Parameter Estimation for Gaussian Mixture and Hidden Markov Models. International Computer Science Institute, Berkeley (1998)
Google Scholar
Jesus, S.C.: A Hybrid System with Symbolic AI and Statistical Methods for Speech Recognition, Doctoral Thesis, University of Washington (1995)
Google Scholar

Download references

Author information

Authors and Affiliations

Computing Research Center, National Polytechnic Institute, Juan de Dios Batiz s/n, P.O. 07038, Mexico
Sergio Suárez Guerra, José Luis Oropeza Rodríguez, Edgardo M. Felipe Riveron & Jesús Figueroa Nazuno

Authors

Sergio Suárez Guerra
View author publications
You can also search for this author in PubMed Google Scholar
José Luis Oropeza Rodríguez
View author publications
You can also search for this author in PubMed Google Scholar
Edgardo M. Felipe Riveron
View author publications
You can also search for this author in PubMed Google Scholar
Jesús Figueroa Nazuno
View author publications
You can also search for this author in PubMed Google Scholar

Editor information

Editors and Affiliations

Dept. System Engineering and Automation, Universitat Politècnica de Catalunya (UPC) Barcelona, Spain
Alberto Sanfeliu
Pattern Recognition Group, ICIMAF, Havana, Cuba
Manuel Lazo Cortés

Rights and permissions

Reprints and permissions

Copyright information

About this paper

Cite this paper

Guerra, S.S., Rodríguez, J.L.O., Riveron, E.M.F., Nazuno, J.F. (2005). Speech Recognition Using Energy Parameters to Classify Syllables in the Spanish Language. In: Sanfeliu, A., Cortés, M.L. (eds) Progress in Pattern Recognition, Image Analysis and Applications. CIARP 2005. Lecture Notes in Computer Science, vol 3773. Springer, Berlin, Heidelberg. https://doi.org/10.1007/11578079_18

Download citation

DOI: https://doi.org/10.1007/11578079_18
Publisher Name: Springer, Berlin, Heidelberg
Print ISBN: 978-3-540-29850-2
Online ISBN: 978-3-540-32242-9
eBook Packages: Computer ScienceComputer Science (R0)

Publish with us

Policies and ethics

Societies and partnerships

The International Association for Pattern Recognition (opens in a new tab)

Speech Recognition Using Energy Parameters to Classify Syllables in the Spanish Language

Abstract

Chapter PDF

Similar content being viewed by others

Automatic Syllabification and Syllable Timing of Automatically Recognized Speech – for Czech

Automatic Syllable Repetition Detection in Continuous Speech Based on Linear Prediction Coefficients

A study on conventional and syllable-based approaches for automatic speech recognition in Malayalam

Keywords

References

Author information

Authors and Affiliations

Editor information

Editors and Affiliations

Rights and permissions

Copyright information

About this paper

Cite this paper

Download citation

Publish with us

Societies and partnerships

Navigation

Speech Recognition Using Energy Parameters to Classify Syllables in the Spanish Language

Abstract

Chapter PDF

Similar content being viewed by others

Automatic Syllabification and Syllable Timing of Automatically Recognized Speech – for Czech

Automatic Syllable Repetition Detection in Continuous Speech Based on Linear Prediction Coefficients

A study on conventional and syllable-based approaches for automatic speech recognition in Malayalam

Keywords

References

Author information

Authors and Affiliations

Editor information

Editors and Affiliations

Rights and permissions

Copyright information

About this paper

Cite this paper

Download citation

Share this paper

Publish with us

Societies and partnerships

Search

Navigation