Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Tigre Broadcast Speech Corpus

Domain:

natural language processing

Record type:

dataset
Creator:
Bei
Host:
A large-scale, open-source speech dataset for the Tigre language (ISO 639-3: tig), developed to support Automatic Speech Recognition (ASR), speech technology research, and language documentation for one of the least-resourced languages in the Afro-Asiatic family.

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

Tigré

Tags

Tigre languageEritreaAudioSpeech CorpusASRlow-resource

Licenses

cc-by-sa-4.0

Similar

Tigre Wikipedia CorpusBeitTigreAI/tigre-speech-text-alignedCommon Voice Scripted Speech 26.0 - TigreSouth African Broadcast News (SABN) CorpusAutomatic Dialect Detection in Arabic Broadcast SpeechGALE Phase 2 Arabic Broadcast Conversation Speech Part 1

Tigre Wikipedia Corpus

This repository houses the Tigre Wikipedia Corpus, a foundational linguistic resource containing all

BeitTigreAI/tigre-speech-text-aligned

This Tigre Speech Corpus is a curated collection of 18,470 aligned audio–text pairs designed to supp

Common Voice Scripted Speech 26.0 - Tigre

A collection of read speech recordings in Tigre (ትግረ).

South African Broadcast News (SABN) Corpus

The corpus consists of approximately 20 hours of audio recordings from one of the count

Automatic Dialect Detection in Arabic Broadcast Speech

We investigate different approaches for dialect identification in Arabic broadcast speech, using pho

GALE Phase 2 Arabic Broadcast Conversation Speech Part 1

Introduction


GALE Phase 2 Arabic Broadcast Conversation Speech Part 1 was developed