Automatic Speech Recognition: An Improved Paradigm

Abstract : In this paper we present a short survey of automatic speech recognition systems underlining the current achievements and capabilities of current day solutions as well as their inherent limitations and shortcomings. In response to which we propose an improved paradigm and algorithm for building an automatic speech recognition system that actively adapts its recognition model in an unsupervised fashion by listening to continuous human speech. The paradigm relies on creating a semi-autonomous system that samples continuous human speech in order to record phonetic units. Then processes those phoneme sized samples to identify the degree of similarity of each sample that will allow the detection of the same phoneme across many samples. After a sufficiently large database of samples has been gathered the system clusters the samples based on their degree of similarity, creating a different cluster for each phoneme. After that the system trains one neural network for each cluster using the samples in that cluster. After a few iterations of sampling, processing, clustering and training the system should contain a neural network detector for each phoneme unit of the spoken language that the system has been exposed to, and be able to use these detectors to recognize phonemes from live speech. Finally we provide the structure and algorithms for this novel automatic speech recognition paradigm.
Type de document :
Communication dans un congrès
Luis M. Camarinha-Matos. 2nd Doctoral Conference on Computing, Electrical and Industrial Systems (DoCEIS), Feb 2011, Costa de Caparica, Portugal. Springer, IFIP Advances in Information and Communication Technology, AICT-349, pp.269-276, 2011, Technological Innovation for Sustainability. 〈10.1007/978-3-642-19170-1_29〉
Liste complète des métadonnées

Littérature citée [22 références]  Voir  Masquer  Télécharger

https://hal.inria.fr/hal-01566588
Contributeur : Hal Ifip <>
Soumis le : vendredi 21 juillet 2017 - 11:25:43
Dernière modification le : vendredi 21 juillet 2017 - 11:30:44

Fichier

978-3-642-19170-1_29_Chapter.p...
Fichiers produits par l'(les) auteur(s)

Licence


Distributed under a Creative Commons Paternité 4.0 International License

Identifiants

Citation

Tudor-Sabin Topoleanu, Gheorghe Mogan. Automatic Speech Recognition: An Improved Paradigm. Luis M. Camarinha-Matos. 2nd Doctoral Conference on Computing, Electrical and Industrial Systems (DoCEIS), Feb 2011, Costa de Caparica, Portugal. Springer, IFIP Advances in Information and Communication Technology, AICT-349, pp.269-276, 2011, Technological Innovation for Sustainability. 〈10.1007/978-3-642-19170-1_29〉. 〈hal-01566588〉

Partager

Métriques

Consultations de la notice

57

Téléchargements de fichiers

36