Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Gauging the accuracy of automatic speech data harvesting in five under-resourced languages

Domaine:

natural language processing
Créateur:
Jaco BadenhorstFebe de Wet
Éditeur:
Dig
Hôte:
Recent research on deep-learning architectures has resulted in substantial improvements in automatic speech recognition accuracy. The leaps of progress made in well-resourced languages can be attributed to the fact that these architectures are able to effectively represent spoken language in all its diversity and complexity. However, developing advanced models of a language without appropriate corpora of speech and text data remains a challenge. For many under-resourced languages, including those spoken in South Africa, such resources simply do not exist. The aim of the work reported on in this paper is to address this situation by investigating the possibility to create diverse speech resources from unannotated broadcast data. The paper describes how existing speech and text resources were used to develop a semi-automatic data harvesting procedure for two genres of broadcast data, namely news bulletins and radio dramas. It was found that adapting acoustic models with less than 10 hours of manually annotated data from the same domain significantly reduced transcription error rates for speaking styles and acoustic conditions that are not represented in any of the existing speech corpora. Results also indicated that much more automatically transcribed adaptation data is required to achieve similar results.

Visit

doi.org

Tasks

automatic speech recognitionspeech processing

Licenses

https://creativecommons.org/licenses/by-sa/4.0

Similaires

Automatic speech recognition for under-resourced languages: A surveyAutomatic speech recognition for an under-resourced language - amharicSpeech recognition for under-resourced languages: Data sharing in hidden Markov model systemsData-Efficient Strategies for Expanding Hate Speech Detection into Under-Resourced LanguagesImproving Speech Recognition for Under-resourced Languages Utilizing Audio-codecs for Data AugmentationEnd-To-End Multilingual Automatic Speech Recognition For Less-Resourced Languages: The Case Of Four Ethiopian Languages

Automatic speech recognition for under-resourced languages: A survey

(Impact-F 1.28 estim. in 2012) International audience no abstract

Automatic speech recognition for an under-resourced language - amharic

Speech recognition for under-resourced languages: Data sharing in hidden Markov model systems

For purposes of automated speech recognition in under-resourced environments, t

Data-Efficient Strategies for Expanding Hate Speech Detection into Under-Resourced Languages

Hate speech is a global phenomenon, but most hate speech datasets so far focus on English-language c

Improving Speech Recognition for Under-resourced Languages Utilizing Audio-codecs for Data Augmentation

Presenter: Nirayo Hailu Gebreegziabher, Ingo Siegert, Andreas Nürnberger, MMSP 2020, Virtual Event,

End-To-End Multilingual Automatic Speech Recognition For Less-Resourced Languages: The Case Of Four Ethiopian Languages

Presenter: Solomon Teferra Abate, Martha Yifiru Tachbelie, Tanja Schultz , ICASSP 20