A study on spoken language identification using deep neural networks

Autor:	Alexandra Draghici, Hanna Lukashevich, Jakob Abeßer
Rok vydání:	2020
Předmět:	Set (abstract data type) Spoken language identification Class (computer programming) Recurrent neural network Computer Science::Sound Computer science Speech recognition Deep neural networks Spectrogram Computer Science::Computation and Language (Computational Linguistics and Natural Language and Speech Processing) Replicate Convolutional neural network
Zdroj:	Audio Mostly Conference
DOI:	10.1145/3411109.3411123
Popis:	In this paper, we investigate a previously proposed algorithm for spoken language identification based on convolutional neural networks and convolutional recurrent neural networks. We improve the algorithm by modifying the training strategy to ensure equal class distribution and efficient memory usage. We successfully replicate previous experimental findings using a modified set of languages. Our findings confirm that both a convolutional neural network as well as convolutional recurrent neural networks are capable to learn language-specific patterns in mel spectrogram representations of speech recordings.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_________::18b7184c23e13ce76d3d9347720d214f https://doi.org/10.1145/3411109.3411123 Zobrazit plný text záznamu