Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

FOOCTTS: Generating Arabic Speech with Acoustic Environment for Football Commentator

Domaine:

natural language processing

Type de record:

papermodelsoftware
Créateur:
BaaAli, Ahmed
Hôte:avatar
This paper presents FOOCTTS, an automatic pipeline for a football commentator that generates speech with background crowd noise. The application gets the text from the user, applies text pre-processing such as vowelization, followed by the commentator's speech synthesizer. Our pipeline included Arabic automatic speech recognition for data labeling, CTC segmentation, transcription vowelization to match speech, and fine-tuning the TTS. Our system is capable of generating speech with its acoustic environment within limited 15 minutes of football commentator recording. Our prototype is generalizable and can be easily applied to different domains and languages. Accepted at Interspeech 2023 Show & Tell Demo Session

Visit

arxiv.org

Tasks

speech processingtext to speech

Tags

Audio and Speech ProcessingComputation and LanguageSound

Similaires

Generating Arabic text in multilingual speech-to-speech machine translation frameworkImproved Speech Pre-Training with Supervision-Enhanced Acoustic UnitArabic Vowels Acoustic CharacterizationGenerating Talking Face Landmarks from SpeechDevelopment of Hausa Acoustic Model for Speech RecognitionFarm Environment Acoustic Monitoring Survey dataset

Generating Arabic text in multilingual speech-to-speech machine translation framework

Improved Speech Pre-Training with Supervision-Enhanced Acoustic Unit

Speech pre-training has shown great success in learning useful and general latent representations fr

Arabic Vowels Acoustic Characterization

International audience A sentence is constructed using basic word units. Each word is

Generating Talking Face Landmarks from Speech

The presence of a corresponding talking face has been shown to significantly improve speech intellig

Development of Hausa Acoustic Model for Speech Recognition

Farm Environment Acoustic Monitoring Survey dataset

This research aims to determine whether farmers monitor the acoustic environment of the