Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

End-To-End Multilingual Automatic Speech Recognition For Less-Resourced Languages: The Case Of Four Ethiopian Languages

Domaine:

natural language processing

Type de record:

paper
Créateur:
SolMarTan
Éditeur:
IEEE
Hôte:avatar
Presenter: Solomon Teferra Abate, Martha Yifiru Tachbelie, Tanja Schultz , ICASSP 2021, Virtual Event, June 6-11, 2021 End-to-End (E2E) approach, which maps a sequence of input features into a sequence of grapheme or words, to Automatic Speech Recognition (ASR) is a hot research agenda. It is interesting for less-resourced languages since it avoids the use of pronunciation dictionary, which is one of the major components in the traditional ASR systems. However, like any deep neural network (DNN) approaches, E2E is data greedy. This makes the application of E2E to less-resourced languages questionable. However, using data from other languages in a multilingual (ML) setup is being applied to solve the problem of data scarcity. We have, therefore, conducted ML E2E ASR experiments for four less-resourced Ethiopian languages using different language and acoustic modelling units. The results of our experiments show that relative Word Error Rate (WER) reductions (over the monolingual E2E systems) of up to 29.83% can be achieved by just using data of two related languages in E2E ASR system training. Moreover, we have also noticed that the use of data from less related languages also leads to E2E ASR performance improvement over the use of monolingual data.

Visit

doi.orgrc.signalprocessingsociety.org

Tasks

automatic speech recognitionspeech processing

Languages

Amharic

Similaires

End-to-End Text-To-Speech synthesis for under resourced South African languagesTowards End-to-End Training of Automatic Speech Recognition for Nigerian PidginDeep Neural Networks Based Automatic Speech Recognition For Four Ethiopian LanguagesMultilingual Speech Recognition With A Single End-To-End ModelSub-word Based End-to-End Speech Recognition for an Under-Resourced Language: AmharicAutomatic speech recognition for under-resourced languages: A survey

End-to-End Text-To-Speech synthesis for under resourced South African languages

Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin

Nigerian Pidgin remains one of the most popular languages in West Africa. With at least 75 million speakers along the West African coast, the language has spread to diasporic communities through Nigerian immigrants in England, Canada, and America, amongst others. I

Deep Neural Networks Based Automatic Speech Recognition For Four Ethiopian Languages

Presenter: Solomon Teferra Abate, ICASSP 2020, Virtual Event, May 4-8, 2020 In this work, we present

Multilingual Speech Recognition With A Single End-To-End Model

Training a conventional automatic speech recognition (ASR) system to support multiple languages is c

Sub-word Based End-to-End Speech Recognition for an Under-Resourced Language: Amharic

In this work, we focused on end-to-end speech recognition for less-resourced language, Amharic. The result can be integrated with other tasks such as spoken content retrieval. We explored three models, which consist of Convolutional Neural Networks, Recurrent Neura

Automatic speech recognition for under-resourced languages: A survey

(Impact-F 1.28 estim. in 2012) International audience no abstract