A performance evaluation of several artificial neural networks for mapping speech spectrum parameters

Yeom Song, Victor; Zeledón Córdoba, Marisol; Coto Jiménez, Marvin

A performance evaluation of several artificial neural networks for mapping speech spectrum parameters

dc.creator	Yeom Song, Victor
dc.creator	Zeledón Córdoba, Marisol
dc.creator	Coto Jiménez, Marvin
dc.date.accessioned	2022-03-24T16:29:55Z
dc.date.available	2022-03-24T16:29:55Z
dc.date.issued	2020
dc.description	Part of the Communications in Computer and Information Science book series (CCIS, volume 1087).	es
dc.description.abstract	In this work, we compare different neural network architectures, for the task of mapping spectral coefficients of noisy speech signals with those corresponding to natural speech. In previous works on the subject, fully-connected multilayer perception (MLP) networks and recurrent neural networks (LSTM & BLSTM) have been used. Several references report some initial trial and error processes to determine which architecture to use. Finding the best network type and size is of great importance due to the considerable training time required by some models of recurrent networks. In our work, we conducted extensive tests training more than five hundred networks, with several architectures to determine which cases present significant differences. The results show that for this application of neural networks, the architectures with more layers or the greater number of neurons are not the most convenient, both for the time required in their training and for the adjustment achieved. These results depend on the complexity of the task (the signal-to-noise ratio or SNR) and the amount of data available. This exploration can guide the most efficient use of these types of neural networks in future mapping applications, and can help to optimize resources in future studies by reducing computational time and complexity.	es
dc.description.procedence	UCR::Vicerrectoría de Docencia::Ingeniería::Facultad de Ingeniería::Escuela de Ingeniería Eléctrica	es
dc.description.sponsorship	Universidad de Costa Rica/[322-B9-105]/UCR/Costa Rica	es
dc.identifier.citation	https://link.springer.com/chapter/10.1007/978-3-030-41005-6_20
dc.identifier.codproyecto	322-B9105
dc.identifier.doi	https://doi.org/10.1007/978-3-030-41005-6_20
dc.identifier.isbn	978-3-030-41005-6
dc.identifier.uri	https://hdl.handle.net/10669/86276
dc.language.iso	eng
dc.source	High Performance Computing (pp.291-306).Turrialba, Costa Rica: Springer, Cham	es
dc.subject	Deep learning	es
dc.subject	Long short-term memory (LSTM)	es
dc.subject	NOISE	es
dc.subject	Speech enhancement	es
dc.title	A performance evaluation of several artificial neural networks for mapping speech spectrum parameters	es
dc.type	comunicación de congreso	es

Files

Original bundle

Now showing 1 - 1 of 1

Name:: Springer5.pdf
Size:: 362.06 KB
Format:: Adobe Portable Document Format
Description:

Download

License bundle

Now showing 1 - 1 of 1

Name:: license.txt
Size:: 3.5 KB
Format:: Item-specific license agreed upon to submission
Description:

Download

Collections

Ingeniería eléctrica