Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

CUPE Universal Phoneme Encoder Substitution for Low-Resource Word Error Rate Reduction in Common Voice

Domaine:

natural language processing

Type de record:

paper
Créateur:
SOV
Éditeur:
Zenodo
Hôte:avatar
Word-piece models (WPMs) are commonly used subword units in state-of-the-art end-to-end automatic speech recognition (ASR) systems. For multilingual ASR, due to the differences in written scripts across languages, multilingual WPMs bring the challenges of having overly large output layers and scaling to more languages. In this work, we propose a universal monolingual output layer (UML) to address such problems. Instead of one output node for only one WPM, UML re-associates each output node with multiple WPMs, one for each language, and results in a smaller monolingual output layer shared acros Research goal: What is the impact of replacing language-specific output layers with the CUPE universal phoneme encoder on word error rate metrics for low-resource languages in the Common Voice dataset? Autonomous synthesis report generated by SOVEREIGN Research Kernel. Tribunal consensus score: 8.0/10. This report was generated autonomously by SOVEREIGN Research Kernel, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 8.0/10.

Visit

doi.orgzenodo.org

Tasks

automatic speech recognitionspeech processing

Tags

impactreplacinglanguage-specificoutputlayersCUPEuniversalphoneme

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode