Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Language ID Prediction from Speech Using Self-Attentive Pooling and 1D-Convolutions

Domaine:

natural language processing

Type de record:

papermodelsoftware
Créateur:
BedMik
Hôte:avatar
This memo describes NTR-TSU submission for SIGTYP 2021 Shared Task on predicting language IDs from speech. Spoken Language Identification (LID) is an important step in a multilingual Automated Speech Recognition (ASR) system pipeline. For many low-resource and endangered languages, only single-speaker recordings may be available, demanding a need for domain and speaker-invariant language ID systems. In this memo, we show that a convolutional neural network with a Self-Attentive Pooling layer shows promising results for the language identification task. Accepted to SYGTYP-2021

Visit

arxiv.org

Tasks

language identificationspeech processing

Tags

Audio and Speech ProcessingComputation and Language