Cargando…

Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks

The speech signal contains a vast spectrum of information about the speaker such as speakers’ gender, age, accent, or health state. In this paper, we explored different approaches to automatic speaker’s gender classification and age estimation system using speech signals. We applied various Deep Neu...

Descripción completa

Detalles Bibliográficos
Autores principales: Kwasny, Damian, Hemmerling, Daria
Formato: Online Artículo Texto
Lenguaje:English
Publicado: MDPI 2021
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC8309811/
https://www.ncbi.nlm.nih.gov/pubmed/34300525
http://dx.doi.org/10.3390/s21144785
_version_ 1783728611069526016
author Kwasny, Damian
Hemmerling, Daria
author_facet Kwasny, Damian
Hemmerling, Daria
author_sort Kwasny, Damian
collection PubMed
description The speech signal contains a vast spectrum of information about the speaker such as speakers’ gender, age, accent, or health state. In this paper, we explored different approaches to automatic speaker’s gender classification and age estimation system using speech signals. We applied various Deep Neural Network-based embedder architectures such as x-vector and d-vector to age estimation and gender classification tasks. Furthermore, we have applied a transfer learning-based training scheme with pre-training the embedder network for a speaker recognition task using the Vox-Celeb1 dataset and then fine-tuning it for the joint age estimation and gender classification task. The best performing system achieves new state-of-the-art results on the age estimation task using popular TIMIT dataset with a mean absolute error (MAE) of 5.12 years for male and 5.29 years for female speakers and a root-mean square error (RMSE) of 7.24 and 8.12 years for male and female speakers, respectively, and an overall gender recognition accuracy of 99.60%.
format Online
Article
Text
id pubmed-8309811
institution National Center for Biotechnology Information
language English
publishDate 2021
publisher MDPI
record_format MEDLINE/PubMed
spelling pubmed-83098112021-07-25 Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks Kwasny, Damian Hemmerling, Daria Sensors (Basel) Article The speech signal contains a vast spectrum of information about the speaker such as speakers’ gender, age, accent, or health state. In this paper, we explored different approaches to automatic speaker’s gender classification and age estimation system using speech signals. We applied various Deep Neural Network-based embedder architectures such as x-vector and d-vector to age estimation and gender classification tasks. Furthermore, we have applied a transfer learning-based training scheme with pre-training the embedder network for a speaker recognition task using the Vox-Celeb1 dataset and then fine-tuning it for the joint age estimation and gender classification task. The best performing system achieves new state-of-the-art results on the age estimation task using popular TIMIT dataset with a mean absolute error (MAE) of 5.12 years for male and 5.29 years for female speakers and a root-mean square error (RMSE) of 7.24 and 8.12 years for male and female speakers, respectively, and an overall gender recognition accuracy of 99.60%. MDPI 2021-07-13 /pmc/articles/PMC8309811/ /pubmed/34300525 http://dx.doi.org/10.3390/s21144785 Text en © 2021 by the authors. https://creativecommons.org/licenses/by/4.0/Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (https://creativecommons.org/licenses/by/4.0/).
spellingShingle Article
Kwasny, Damian
Hemmerling, Daria
Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks
title Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks
title_full Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks
title_fullStr Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks
title_full_unstemmed Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks
title_short Gender and Age Estimation Methods Based on Speech Using Deep Neural Networks
title_sort gender and age estimation methods based on speech using deep neural networks
topic Article
url https://www.ncbi.nlm.nih.gov/pmc/articles/PMC8309811/
https://www.ncbi.nlm.nih.gov/pubmed/34300525
http://dx.doi.org/10.3390/s21144785
work_keys_str_mv AT kwasnydamian genderandageestimationmethodsbasedonspeechusingdeepneuralnetworks
AT hemmerlingdaria genderandageestimationmethodsbasedonspeechusingdeepneuralnetworks