Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Massively Multilingual ASR: 50 Languages, 1 Model, 1 Billion Parameters

Domaine:

natural language processing

Type de record:

papermodel
Créateur:
PraSriTomHan
Hôte:avatar
We study training a single acoustic model for multiple languages with the aim of improving automatic speech recognition (ASR) performance on low-resource languages, and over-all simplifying deployment of ASR systems that support diverse languages. We perform an extensive benchmark on 51 languages, with varying amount of training data by language(from 100 hours to 1100 hours). We compare three variants of multilingual training from a single joint model without knowing the input language, to using this information, to multiple heads (one per language cluster). We show that multilingual training of ASR models on several languages can improve recognition performance, in particular, on low resource languages. We see 20.9%, 23% and 28.8% average WER relative reduction compared to monolingual baselines on joint model, joint model with language input and multi head model respectively. To our knowledge, this is the first work studying multilingual ASR at massive scale, with more than 50 languages and more than 16,000 hours of audio across them.

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Tags

Audio and Speech ProcessingComputation and LanguageSound

Similaires

Massively Multilingual Text Translation For Low-Resource LanguagesGemDetox at TextDetox CLEF 2025: Enhancing a Massively Multilingual Model for Text Detoxification on Low-resource LanguagesMassively Multilingual Word EmbeddingsLearning ASR pathways: A sparse multilingual ASR model75 Languages, 1 Model: Parsing Universal Dependencies UniversallyMultilingual Open Text Release 1: Public Domain News in 44 Languages

Massively Multilingual Text Translation For Low-Resource Languages

Translation into severely low-resource languages has both the cultural goal of saving and reviving t

GemDetox at TextDetox CLEF 2025: Enhancing a Massively Multilingual Model for Text Detoxification on Low-resource Languages

As social-media platforms emerge and evolve faster than the regulations meant to oversee them, autom

Massively Multilingual Word Embeddings

We introduce new methods for estimating and evaluating embeddings of words in more than fifty langua

Learning ASR pathways: A sparse multilingual ASR model

Neural network pruning compresses automatic speech recognition (ASR) models effectively. However, in

75 Languages, 1 Model: Parsing Universal Dependencies Universally

We present UDify, a multilingual multi-task model capable of accurately predicting universal part-of

Multilingual Open Text Release 1: Public Domain News in 44 Languages

We present Multilingual Open Text (MOT), a new multilingual corpus containing text in 44 languages,