Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

FFSTC: Fongbe to French Speech Translation Corpus

Domain:

natural language processing

Record type:

paperdataset
Creator:
KpoLalEzi
Host:avatar
In this paper, we introduce the Fongbe to French Speech Translation Corpus (FFSTC) for the first time. This corpus encompasses approximately 31 hours of collected Fongbe language content, featuring both French transcriptions and corresponding Fongbe voice recordings. FFSTC represents a comprehensive dataset compiled through various collection methods and the efforts of dedicated individuals. Furthermore, we conduct baseline experiments using Fairseq's transformer_s and conformer models to evaluate data quality and validity. Our results indicate a score of 8.96 for the transformer_s model and 8.14 for the conformer model, establishing a baseline for the FFSTC corpus.

Visit

arxiv.org

Tasks

machine translationspeech processingspeech translation

Languages

Fon

Tags

Computation and Language

Similar

FFSTC 2: Extending the Fongbe to French Speech Translation CorpusExtending the Fongbe to French Speech Translation Corpus: resources, models and benchmarkFrench-Fongbe Parallel CorpusImproving End-to-End Speech Translation for the Low Resource Language Fongbe to French

FFSTC 2: Extending the Fongbe to French Speech Translation Corpus

Extending the Fongbe to French Speech Translation Corpus: resources, models and benchmark

French-Fongbe Parallel Corpus

Ce dataset est un corpus parallèle Français-Fongbe (Bénin) généré par IA et structuré pour l'entraîn

Improving End-to-End Speech Translation for the Low Resource Language Fongbe to French

This study addresses the challenges of end-to-end (E2E) Speech-to-Text Translation (STT) for the low