Λεπτομέρειες βιβλιογραφικής εγγραφής
| Τίτλος: |
Enhancing Korean-Accented English ASR with Transliteration-Based Data Synthesis. |
| Συγγραφείς: |
Jang, Hana, Kim, Taehwa, Choi, Hyungwoo, Jung, Youngbeom |
| Πηγή: |
Electronics (2079-9292); Apr2026, Vol. 15 Issue 7, p1380, 23p |
| Θεματικοί όροι: |
Transliteration, Data augmentation, Automatic speech recognition, Error rates, Korean language, Phonetic transcriptions |
| Περίληψη: |
Despite recent advances in automatic speech recognition (ASR), performance remains limited for Korean-accented English due to the limited availability of accent-specific speech data, including pronunciation and prosodic variations. To address this limitation, we propose a synthetic data generation framework for improving Whisper-based ASR performance. Synthetic speech is generated by converting English text into Hangul-based phonetic transcriptions using an intermediate IPA representation to reflect the phonological characteristics of Korean-accented English. The ASR model is fine-tuned using Low-Rank Adaptation with a mixture of synthetic and authentic speech data. Experimental results demonstrate relative reductions of up to 16.40% in the character error rate, 14.93% in the word error rate, and 14.81% in the phoneme error rate compared to the pretrained baseline. [ABSTRACT FROM AUTHOR] |
|
Copyright of Electronics (2079-9292) is the property of MDPI and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.) |
| Βάση Δεδομένων: |
Complementary Index |