Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Spatial models of linguistic data in Africa and beyond

Domaine:

natural language processing

Type de record:

paper
Créateur:
Mertner, Miri
Éditeur:
UniUniJäg
Éditeur:
Uni
Hôte:avatar
The spatial distribution of linguistic diversity and structural elements of language, like morphosyntactic and phonological features, is a rich source of knowledge for those interested in historical linguistics and the evolution of language. Bayesian spatial models are a promising way to uncover these patterns. However, modelling linguistic data in space comes with a unique set of complexities and considerations. These include questions of how to represent the geographic locations of languages, how to measure the distance between them, and how to account for variations in topography. In this thesis, I present a series of case studies on African languages which will illustrate different Bayesian spatial models, all of which simultaneously incorporate information about language history in the form of phylogenies or language families. Recognising that the application of spatial models in linguistic typology is a relatively recent development, I will provide a general overview of some common models used in spatial statistics and describe the approaches which have been used in this thesis. Environmental and social factors have long been thought to influence the distribution of linguistic diversity in time and space. I will use a novel combination of methods to show that the factors which impact recent diversification are distinct from those which impact the maintenance of diversity over time. Following that, I will examine areal patterns of structural convergence between African languages and uncover systematic variation in the diffusibility of structural elements of language. Lastly, I will examine geographic patterns of data sparsity and discuss their implications for future statistical studies in linguistics.

Visit

doi.orgpublikationen.uni-tuebingen.de

Tags

linguisticsafrican languagesspatial modelsbayesianstatistics

Licenses

ubt-podnohttp://tobias-lib.uni-tuebingen.de/doku/lic_ohne_pod.php?la=dehttp://tobias-lib.uni-tuebingen.de/doku/lic_ohne_pod.php?la=en

Similaires

Beyond data gaps: tracking spatial inequality in Africa via nighttime lightsBeyond Swadesh Beyond Swadesh: Recalibrating Core Wordlists for Detecting Deep Linguistic Relationships in West AfricaBeyond Data Quantity: Key Factors Driving Performance in Multilingual Language ModelsSpatial models of animal mobility: buffalo and cattle in Southern AfricaColour in Translation: Data, Models, and Benchmarking for Cross-Linguistic Colour NamingTemporal and Spatial Data Mining with Second-Order Hidden Markov Models

Beyond data gaps: tracking spatial inequality in Africa via nighttime lights

Abstract This paper offers a novel and robust approach to measure the distributi

Beyond Swadesh Beyond Swadesh: Recalibrating Core Wordlists for Detecting Deep Linguistic Relationships in West Africa

International audience The classification of the language isolate Bangime hinges on t

Beyond Data Quantity: Key Factors Driving Performance in Multilingual Language Models

Multilingual language models (MLLMs) are crucial for handling text across various languages, yet the

Spatial models of animal mobility: buffalo and cattle in Southern Africa

A spatial model of animal mobility has been developed for buffalo and cattle at the interface betwee

Colour in Translation: Data, Models, and Benchmarking for Cross-Linguistic Colour Naming

Colour naming links vision and language. Yet, effective cross-linguistic colour communication is lim

Temporal and Spatial Data Mining with Second-Order Hidden Markov Models

Colloque avec actes sans comité de lecture. internationale. International audience In