Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

KALAKA-3: a database for the assessment of spoken language recognition technology on YouTube audios

Creator:
LuiMikAmpMir
Publisher:
Spr
Host:

Visit

doi.org

Tasks

language identification

Languages

Kalanga

Licenses

http://www.springer.com/tdm

Similar

VOXLINGUA107: A DATASET FOR SPOKEN LANGUAGE RECOGNITIONCALYOU: A Comparable Spoken Algerian Corpus Harvested from YouTubeA database for Amazigh speech recognition research: AMZSRDOn the Use of Machine Translation for Spoken Language Understanding PortabilityAMHCD: A Database for Amazigh Handwritten Character Recognition Researcholeteane/Database-for-setswana-speech-recognition

VOXLINGUA107: A DATASET FOR SPOKEN LANGUAGE RECOGNITION

This paper investigates the use of automatically collected web audio data for the task of spoken language recognition. We generate semirandom search phrases from language-specific Wikipedia data that are then used to retrieve videos from YouTube for 107 languages.

CALYOU: A Comparable Spoken Algerian Corpus Harvested from YouTube

This paper addresses the issue of comparability of comments extracted from Youtube. The comments concern spoken Algerian that could be either local Arabic, Modern Standard Arabic or French.

This diversity of expression gives rise to a huge number of prob

A database for Amazigh speech recognition research: AMZSRD

On the Use of Machine Translation for Spoken Language Understanding Portability

International audience Across language portability of a spoken language understanding

AMHCD: A Database for Amazigh Handwritten Character Recognition Research

oleteane/Database-for-setswana-speech-recognition

# Database-for-setswana-speech-recognition