☆ 3.8 Proceedings Paper

IMPROVING MANDARIN TONE MISPRONUNCIATION DETECTION FOR NON-NATIVE LEARNERS WITH SOFT-TARGET TONE LABELS AND BLSTM-BASED DEEP MODELS

2018 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP) (2018)

期刊

2018 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP)

卷 -, 期 -, 页码 6249-6253

出版社

IEEE

关键词

Computer assistant language learning (CALL); computer assisted pronunciation training (CAPT); tone recognition and mispronunciation detection; deep learning

类别

Acoustics Engineering, Electrical & Electronic

资金

China Scholarship Council
NFR AULUS project

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

We propose three techniques to improve mispronunciation detection of Mandarin tones of second language (L2) learners using tone-based extended recognition network (ERN). First, we extend our model from deep neural network (DNN) to bidirectional long-short-term memory (BLSTM) in order to model tone-level co-articulation influenced by a broader temporal context (e.g., two or three consecutive Mandarin syllables). Second, we relax the hard labels to characterize the situations when a single tone class label is not enough because L2 learners' pronunciations are often between two canonical tone categories. Therefore, soft targets (a probabilistic transcription) are proposed for acoustic model training in place of conventional hard targets (one-hot targets). Third, we average tone scores produced by BLSTM models trained with hard and soft targets to seek the complementarity from modeling at the tone-target levels. Compared to our previous system based on the DNN-trained ERNs, the BLSTM-trained system with soft targets reduces the equal error rate (ERR) from 5.77% to 4.86%, and system combination decreases EER further to 4.34%, achieving a 24.78% relative error reduction.

IMPROVING MANDARIN TONE MISPRONUNCIATION DETECTION FOR NON-NATIVE LEARNERS WITH SOFT-TARGET TONE LABELS AND BLSTM-BASED DEEP MODELS

期刊

2018 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP)

出版社

IEEE

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

IMPROVING MANDARIN TONE MISPRONUNCIATION DETECTION FOR NON-NATIVE LEARNERS WITH SOFT-TARGET TONE LABELS AND BLSTM-BASED DEEP MODELS

期刊

2018 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP)

出版社

IEEE

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文