Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

VoxLingua107 ECAPA-TDNN Spoken Language Identification Model

Domaine:

natural language processing

Type de record:

model
This is a spoken language recognition model trained on the VoxLingua107 dataset using SpeechBrain. The model uses the ECAPA-TDNN architecture that has previously been used for speaker recognition. The model can classify a speech utterance according to the language spoken. It covers 107 different languages

Visit

huggingface.co

Connected records

paperdataset

Tasks

keywordsautomatic speech recognitionspeech processinglanguage identification

Languages

AfrikaansAmharicHausaLingalaMalagasyShonaSomaliSwahiliYoruba

Tags

VoxLingua107

Similaires

VOXLINGUA107: A DATASET FOR SPOKEN LANGUAGE RECOGNITIONtsuxalo/Spoken-Language-Translation-ModelVoxLingua107Development of a Spoken Language Identification System for South African LanguagesSpoken Arabic Algerian dialect identificationLSTM-TDNN with convolutional front-end for Dialect Identification in the 2019 Multi-Genre Broadcast Challenge

VOXLINGUA107: A DATASET FOR SPOKEN LANGUAGE RECOGNITION

This paper investigates the use of automatically collected web audio data for the task of spoken language recognition. We generate semirandom search phrases from language-specific Wikipedia data that are then used to retrieve videos from YouTube for 107 languages.

tsuxalo/Spoken-Language-Translation-Model

A Framework for Translating Hausa Audio Recordings into English Text # Spoken-Language-Translation-

VoxLingua107

VoxLingua107 is a speech dataset for training spoken language identification models. The dataset consists of short speech segments automatically extracted from YouTube videos and labeled according the language of the video title and description, with some post-proc

Development of a Spoken Language Identification System for South African Languages

Spoken Arabic Algerian dialect identification

LSTM-TDNN with convolutional front-end for Dialect Identification in the 2019 Multi-Genre Broadcast Challenge

This paper presents a novel Dialect Identification (DID) system developed for the Fifth Edition of t