상세 보기
Automatic detection of speech sound disorder in children using automatic speech recognition and audio classification
- Selina S. Sung;
- Jungmin So;
- Tae-Jin Yoon;
- Seunghee Ha
초록
Children with speech sound disorders (SSDs) face various challenges in producing speech sounds, which often lead to significant social and educational barriers. Detecting and treating SSDs in children is complex due to the variability in disorder severity and diagnostic boundaries. This study aims to develop an automated SSD detection system using deep learning models, leveraging their ability to transcribe audio, efficiently capture sound patterns on a vast scale, and address the limitations of traditional methods involving speech-language pathologists. For this study, we collected audio recordings from 573 children aged two to nine using standardized prompts from the Assessment of Phonology and Articulation for Children. Speech-language pathologists analyzed the recordings and identified 92 children with SSDs. To build an automatic SSD detection system, we used a dataset to train neural network models for automatic speech recognition and audio classification. Five different methods are studied, with the best method achieving 73.9% unweighted average recall. While the results show the potential of using deep learning models for the automatic detection of SSDs in children, further research is needed to improve the reliability of the models widely used in practice.
키워드
- 제목
- Automatic detection of speech sound disorder in children using automatic speech recognition and audio classification
- 저자
- Selina S. Sung; Jungmin So; Tae-Jin Yoon; Seunghee Ha
- 발행일
- 2024-09
- 저널명
- 말소리와 음성과학
- 권
- 16
- 호
- 3
- 페이지
- 87 ~ 94