RESEARCH
Sunhee Kim, Jooyoung Lee, Seo Gyeong Choi, Seunghun Ji, Jeemin Kang, Jongin Kim, Dohee Kim, Boryong Kim, Eungi Cho, Hojeong Kim, Jeongmin Jang, Jun Hyung Kim, Bon Hyeok Ku, Hyungmin Park, Minhwa Chung
This paper describes a method of building Korean conversational speech data in the emergency medical domain and proposes an annotation method for the collected data in order to improve speech recognition performance. To suggest future research directions, baseline speech recognition experiments were conducted by using partial data that were collected and annotated. All voices were recorded at 16-bit resolution at 16 kHz sampling rate. A total of 166 conversations were collected, amounting to 8 hours and 35 minutes. Various information was manually transcribed such as orthography, pronunciation, dialect, noise, and medical information using Praat. Baseline speech recognition experiments were used to depict problems related to speech recognition in the emergency medical domain. The Korean conversational speech data presented in this paper are first-stage data in the emergency medical domain and are expected to be used as training data for developing conversational systems for emergency medical applications.
No. | Title. | Date. |
---|