Automatic detection of speech sound disorder in children using automatic speech recognition and audio classification

초록

Children with speech sound disorders (SSDs) face various challenges in producing speech sounds, which often lead to significant social and educational barriers. Detecting and treating SSDs in children is complex due to the variability in disorder severity and diagnostic boundaries. This study aims to develop an automated SSD detection system using deep learning models, leveraging their ability to transcribe audio, efficiently capture sound patterns on a vast scale, and address the limitations of traditional methods involving speech-language pathologists. For this study, we collected audio recordings from 573 children aged two to nine using standardized prompts from the Assessment of Phonology and Articulation for Children. Speech-language pathologists analyzed the recordings and identified 92 children with SSDs. To build an automatic SSD detection system, we used a dataset to train neural network models for automatic speech recognition and audio classification. Five different methods are studied, with the best method achieving 73.9% unweighted average recall. While the results show the potential of using deep learning models for the automatic detection of SSDs in children, further research is needed to improve the reliability of the models widely used in practice.

키워드

speech sound disorderautomatic speech recognitionaudio classification
제목
Automatic detection of speech sound disorder in children using automatic speech recognition and audio classification
저자
Selina S. SungJungmin SoTae-Jin YoonSeunghee Ha
DOI
10.13064/KSSS.2024.16.3.087
발행일
2024-09
저널명
말소리와 음성과학
16
3
페이지
87 ~ 94