Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Odyssey SA Voice Corpus (V0.4) — Evaluation Preview

Domain:

natural language processing

Record type:

dataset
Creator:
ODY
Host:
Odyssey SA Voice Corpus (V0.4) — Evaluation Preview Overview The Odyssey SA Voice Corpus (V0.4) is a 1000-hour multilingual South African speech dataset spanning 8 languages, designed for automatic speech recognition (ASR), code-switching research, and large audio model (LAM) evaluation.

Visit

huggingface.co

Tasks

automatic speech recognitioncode switchingspeech processing

Languages

AfrikaansNdebeleSetswanaXhosaZulu

Tags

multilingualsouth-africacode-switchingspeechaudioasrevaluation

Licenses

cc-by-4.0

Similar

Odyssey SA Voice Corpus v0.1ODYSSEYAILABS/odyssey-sa-voice-corpus-v0.3ODYSSEYAILABS/odyssey-v0.5.1-16h-eval-preview-public

Odyssey SA Voice Corpus v0.1

language: - en - zu - xh - set - afr license: cc-by-4.0 task_categories: - automatic-speech-recognit

ODYSSEYAILABS/odyssey-sa-voice-corpus-v0.3

license: cc-by-4.0 task_categories: automatic-speech-recognition text-to-speech language: af nr en

ODYSSEYAILABS/odyssey-v0.5.1-16h-eval-preview-public

Odyssey AI Labs presents a 16-hour Enterprise Evaluation Preview of the V0.5 corpus. This release is