A performance evaluation of several artificial neural networks for mapping speech spectrum parameters
dc.creator | Yeom Song, Victor | |
dc.creator | Zeledón Córdoba, Marisol | |
dc.creator | Coto Jiménez, Marvin | |
dc.date.accessioned | 2022-03-24T16:29:55Z | |
dc.date.available | 2022-03-24T16:29:55Z | |
dc.date.issued | 2020 | |
dc.description | Part of the Communications in Computer and Information Science book series (CCIS, volume 1087). | es_ES |
dc.description.abstract | In this work, we compare different neural network architectures, for the task of mapping spectral coefficients of noisy speech signals with those corresponding to natural speech. In previous works on the subject, fully-connected multilayer perception (MLP) networks and recurrent neural networks (LSTM & BLSTM) have been used. Several references report some initial trial and error processes to determine which architecture to use. Finding the best network type and size is of great importance due to the considerable training time required by some models of recurrent networks. In our work, we conducted extensive tests training more than five hundred networks, with several architectures to determine which cases present significant differences. The results show that for this application of neural networks, the architectures with more layers or the greater number of neurons are not the most convenient, both for the time required in their training and for the adjustment achieved. These results depend on the complexity of the task (the signal-to-noise ratio or SNR) and the amount of data available. This exploration can guide the most efficient use of these types of neural networks in future mapping applications, and can help to optimize resources in future studies by reducing computational time and complexity. | es_ES |
dc.description.procedence | UCR::Vicerrectoría de Docencia::Ingeniería::Facultad de Ingeniería::Escuela de Ingeniería Eléctrica | es_ES |
dc.description.sponsorship | Universidad de Costa Rica/[322-B9-105]/UCR/Costa Rica | es_ES |
dc.identifier.citation | https://link.springer.com/chapter/10.1007/978-3-030-41005-6_20 | es_ES |
dc.identifier.codproyecto | 322-B9-105 | |
dc.identifier.doi | 10.1007/978-3-030-41005-6_20 | |
dc.identifier.isbn | 978-3-030-41005-6 | |
dc.identifier.uri | https://hdl.handle.net/10669/86276 | |
dc.language.iso | eng | es_ES |
dc.source | High Performance Computing (pp.291-306).Turrialba, Costa Rica: Springer, Cham | es_ES |
dc.subject | Deep learning | es_ES |
dc.subject | Long short-term memory (LSTM) | es_ES |
dc.subject | NOISE | es_ES |
dc.subject | Speech enhancement | es_ES |
dc.title | A performance evaluation of several artificial neural networks for mapping speech spectrum parameters | es_ES |
dc.type | comunicación de congreso | es_ES |