Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Language ID Prediction from Speech Using Self-Attentive Pooling and 1D-Convolutions

Domain:

natural language processing

Record type:

papermodelsoftware
Creator:
BedMik
Host:avatar
This memo describes NTR-TSU submission for SIGTYP 2021 Shared Task on predicting language IDs from speech. Spoken Language Identification (LID) is an important step in a multilingual Automated Speech Recognition (ASR) system pipeline. For many low-resource and endangered languages, only single-speaker recordings may be available, demanding a need for domain and speaker-invariant language ID systems. In this memo, we show that a convolutional neural network with a Self-Attentive Pooling layer shows promising results for the language identification task. Accepted to SYGTYP-2021

Visit

arxiv.org

Tasks

language identificationspeech processing

Tags

Audio and Speech ProcessingComputation and Language

Similar

Amazigh Speech Recognition Using 1D CNNAutomatic Speech Recognition for Amharic Language using Self-SupervisedGhanaNLP/ghana-speech-idAfriSpeech/african-speech-idLanguage-Based Image Editing with Recurrent Attentive ModelsAn automatic speech recognition system for isolated Amazigh word using 1D & 2D CNN-LSTM architecture

Amazigh Speech Recognition Using 1D CNN

Automatic Speech Recognition for Amharic Language using Self-Supervised

Automatic Speech Recognition (ASR) systems have become a very natural human-machine interaction in w

GhanaNLP/ghana-speech-id

Language identification for Ghanaian and West African speech, over IPA phonemes # ghana-speech-id

AfriSpeech/african-speech-id

CPU-friendly, fast language identification for 1,386 African languages # african-speech-id CPU-fri

Language-Based Image Editing with Recurrent Attentive Models

We investigate the problem of Language-Based Image Editing (LBIE). Given a source image and a natura

An automatic speech recognition system for isolated Amazigh word using 1D & 2D CNN-LSTM architecture