Artículo
Alzheimer disease recognition using speech-based embeddings from pre-trained models
Fecha de publicación:
09/2021
Editorial:
International Speech Communication Association
Revista:
Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH
ISSN:
2308-457X
e-ISSN:
1990-9772
Idioma:
Inglés
Tipo de recurso:
Artículo publicado
Clasificación temática:
Resumen
This paper describes our submission to the ADreSSo Challenge, which focuses on the problem of automatic recognition of Alzheimer's Disease (AD) from speech. The audio samples contain speech from the subjects describing a picture with the guidance of an experimenter. Our approach to the problem is based on the use of embeddings extracted from different pretrained models - trill, allosaurus, and wav2vec 2.0 - which were trained to solve different speech tasks. These features are modeled with a neural network that takes short segments of speech as input, generating an AD score per segment. The final score for an audio file is given by the average over all segments in the file. We include ablation results to show the performance of different feature types individually and in combination, a study of the effect of the segment size, and an analysis of statistical significance. Our results on the test data for the challenge reach an accuracy of 78.9%, outperforming both the acoustic and linguistic baselines provided by the organizers.
Archivos asociados
Licencia
Identificadores
Colecciones
Articulos(ICC)
Articulos de INSTITUTO DE INVESTIGACION EN CIENCIAS DE LA COMPUTACION
Articulos de INSTITUTO DE INVESTIGACION EN CIENCIAS DE LA COMPUTACION
Citación
Gauder, María Lara; Pepino, Leonardo Daniel; Ferrer, Luciana; Riera, Pablo; Alzheimer disease recognition using speech-based embeddings from pre-trained models; International Speech Communication Association; Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH; 6; 9-2021; 4186-4190
Compartir
Altmétricas