Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Mispronunciation Detection in Non-native (L2) English with Uncertainty Modeling

Domain:

natural language processing

Record type:

paper
Creator:
KorLorZapCal
Host:avatar
A common approach to the automatic detection of mispronunciation in language learning is to recognize the phonemes produced by a student and compare it to the expected pronunciation of a native speaker. This approach makes two simplifying assumptions: a) phonemes can be recognized from speech with high accuracy, b) there is a single correct way for a sentence to be pronounced. These assumptions do not always hold, which can result in a significant amount of false mispronunciation alarms. We propose a novel approach to overcome this problem based on two principles: a) taking into account uncertainty in the automatic phoneme recognition step, b) accounting for the fact that there may be multiple valid pronunciations. We evaluate the model on non-native (L2) English speech of German, Italian and Polish speakers, where it is shown to increase the precision of detecting mispronunciations by up to 18% (relative) compared to the common approach. Accepted to ICASSP 2021

Visit

arxiv.org

Tags

Audio and Speech ProcessingMachine LearningSound

Similar

L2-KSU Native and Non-Native Arabic SpeechMispronunciation Detection Without L2 Pronunciation Dataset in Low-Resource Setting: A Case Study in Finland SwedishDuration as a cue to stress and accent in tunisian Arabic, native English, and L2 EnglishThe role of vowel quality in cuing stress and accent in tunisian Arabic, native English, and L2 EnglishCorpus linguistics and non‐native varieties of EnglishThe pitch range of L2 English read by native speakers of Jordanian Arabic compared with that of L1 speakers of English and Arabic

L2-KSU Native and Non-Native Arabic Speech

Introduction

L2-KSU Native and Non-Native Arabic Speech was developed by

Mispronunciation Detection Without L2 Pronunciation Dataset in Low-Resource Setting: A Case Study in Finland Swedish

Mispronunciation detection (MD) models are the cornerstones of many language learning applications.

Duration as a cue to stress and accent in tunisian Arabic, native English, and L2 English

The role of vowel quality in cuing stress and accent in tunisian Arabic, native English, and L2 English

Corpus linguistics and non‐native varieties of English

ABSTRACT: This article derives from the internal discussions of a project that has just been launche

The pitch range of L2 English read by native speakers of Jordanian Arabic compared with that of L1 speakers of English and Arabic

International audience The pitch range of L2 English, produced by native speakers of