Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Statistical modelling of speech units in HMM-based speech synthesis for Arabic

Domaine:

natural language processing

Type de record:

paper
Créateur:
HouColMnaJou
Éditeur:
SpeNat
Éditeur:
CCSD
Hôte:avatar
International audience This paper investigates statistical parametric speech synthesis of Modern Standard Arabic (MSA). Hidden Markov Models (HMM)-based speech synthesis system relies on a description of speech segments corresponding to phonemes, with a large set of features that represent phonetic, phonologic, linguistic and contextual aspects. When applied to MSA two specific phenomena have to be taken in account, the vowel lengthening and the consonant gemination. This paper studies thoroughly the modeling of these phenomena through various approaches: as for example, the use of different units for modeling short vs. long vowels and the use of different units for modeling simple vs. geminated consonants. These approaches are compared to another one which merges short and long variants of a vowel into a single unit and, simple and geminated variants of a consonant into a single unit (these characteristics being handled through the features associated to the sound). Results of subjective evaluation show that there is no significant difference between using the same unit for simple and geminated consonant (as well as for short and long vowels) and using different units for simple vs. geminated consonants (as well for short vs. long vowels).

Visit

inria.hal.science

Tasks

speech processingtext to speech

Tags

speech unit modelingArabic languagestatistical modelingspeech synthesis[INFO.INFO-TS]Computer Science [cs]/Signal and Image Processing

Licenses

https://about.hal.science/hal-authorisation-v1/info:eu-repo/semantics/OpenAccess

Similaires

Statistical parametric speech synthesis for IbibioF0 Modeling In Hmm-Based Speech Synthesis System Using Deep Belief NetworkTone modelling in Ibibio speech synthesisF0 modeling using DNN for Arabic parametric speech synthesisPathological Detection Using HMM Speech Recognition-Based Amazigh DigitsSyllabification for Afrikaans speech synthesis

Statistical parametric speech synthesis for Ibibio

F0 Modeling In Hmm-Based Speech Synthesis System Using Deep Belief Network

In recent years multilayer perceptrons (MLPs) with many hid- den layers Deep Neural Network (DNN) ha

Tone modelling in Ibibio speech synthesis

F0 modeling using DNN for Arabic parametric speech synthesis

International audience Deep neural networks (DNN) are gaining increasing interest in

Pathological Detection Using HMM Speech Recognition-Based Amazigh Digits

Syllabification for Afrikaans speech synthesis