Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Quantifying lexical and pronunciation variation between three Arabic varieties*

Domaine:

natural language processing

Type de record:

paper
Créateur:
MahEla
Éditeur:
Joh
Hôte:
This paper reports on computational measures of linguistic variation that quantify the lexical and pronunciation variation between three varieties of Arabic, Moroccan Arabic, Egyptian Arabic, and Gulf Arabic. We provide three measures of linguistic variation; all computed based on elicitation of the Swadesh list. The first measure is the lexical variation based on the percentage of noncognate words. The second is another lexical measure that takes into account a pronunciation aspect by considering the IPA transcription of the same word list. The third is a pronunciation measure that computes the variation of the IPA transcription of the cognate words in the Swadesh list. The results of the three measures show that geographically proximate languages are also linguistically closer to each other.

Visit

doi.org

Languages

Arabic, Moroccan Spoken

Similaires

Pronunciation Variation: Coloured AfrikaansGlobalPhone Arabic Pronunciation DictionaryAspects of Pronunciation in Five Varieties of EnglishQuantifying Language Variation Acoustically with Few ResourcesTokenization of Tunisian Arabic: A Comparison between Three Machine Learning ModelsNadiaBMKarmani/Tunisian-Arabic-Lexical-Dictionary

Pronunciation Variation: Coloured Afrikaans

Audio recordings and orthographic and phonetic transcriptions of 5 young Coloured speakers in 3 regi

GlobalPhone Arabic Pronunciation Dictionary

The GlobalPhone pronunciation dictionaries, created within the framework of the multilingual speech

Aspects of Pronunciation in Five Varieties of English

The English language is one of the most widely spoken languages in the world, being spoken by approx

Quantifying Language Variation Acoustically with Few Resources

Deep acoustic models represent linguistic information based on massive amounts of data. Unfortunatel

Tokenization of Tunisian Arabic: A Comparison between Three Machine Learning Models

Tokenization represents the way of segmenting a piece of text into smaller units called tokens. Sinc

NadiaBMKarmani/Tunisian-Arabic-Lexical-Dictionary

Main TA clitics.