응급의료 영역 한국어 음성대화 데이터베이스 구축

Building a Korean conversational speech database in the emergency medical domain
  • 김선희
  • 이주영
  • 최서경
  • 지승훈
  • 강지민
  • ... 박형민
  • 외 9명

초록

This paper describes a method of building Korean conversational speech data in the emergency medical domain and proposes an annotation method for the collected data in order to improve speech recognition performance. To suggest future research directions, baseline speech recognition experiments were conducted by using partial data that were collected and annotated. All voices were recorded at 16-bit resolution at 16 kHz sampling rate. A total of 166 conversations were collected, amounting to 8 hours and 35 minutes. Various information was manually transcribed such as orthography, pronunciation, dialect, noise, and medical information using Praat. Baseline speech recognition experiments were used to depict problems related to speech recognition in the emergency medical domain. The Korean conversational speech data presented in this paper are first-stage data in the emergency medical domain and are expected to be used as training data for developing conversational systems for emergency medical applications.

키워드

conversational speechspeech dataspeech recognitionannotationemergency medical domain
제목
응급의료 영역 한국어 음성대화 데이터베이스 구축
제목 (타언어)
Building a Korean conversational speech database in the emergency medical domain
저자
김선희이주영최서경지승훈강지민김종인김도희김보령조은기김호정장정민김준형구본혁박형민정민화
DOI
10.13064/KSSS.2020.12.4.081
발행일
2020-12
저널명
말소리와 음성과학
12
4
페이지
81 ~ 90