Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Language Ecology and Endangerment Dataset for South-South Nigeria

Domaine:

natural language processing
Créateur:
EkpUdoUruUdo
Éditeur:
Uni
Éditeur:
Men
Hôte:avatar
This dataset was obtained from a field survey conducted on proper households (those with at least one parent and one child) in the South-South Geopolitical Zone of Nigeria. The dataset, collected between July 2023 and April 2024 using a purposeful sampling method, includes 543 validated responses captured in real-time through an online, electronic survey (e-survey) instrument developed with Google Forms. The survey instrument was synthesised from the UNESCO 2003 Language Vitality and Endangerment (LVE) framework/questionnaire to capture personalised views from households (five per Local Government Area (LGA)) within the language communities. The synthesised instrument makes the dataset suitable for identifying the causal LVE factors, group(s) or agent(s), thereby supporting efficient knowledge extraction and localisation. Also included are data on the speech systems of the languages spoken in these communities–consisting of recorded speech of 108 words selected from the Swadesh wordlist, with textual documentation of the gloss, syllable, and tone patterns of each word. The dataset is useful for mining insights and identifying patterns in LVE data. It can also assist researchers understand trends, linguistic changes, and interactions over time, and can support the analysis of complex linguistic behaviours across language communities.

Visit

doi.orgdata.mendeley.com

Tags

Arts and HumanitiesLinguisticsFOS: Languages and literatureData MiningGeographic Information System

Licenses

info:eu-repo/semantics/openAccessCreative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

A multimodal dataset for automating language vitality and endangerment assessment in south-south NigeriaA novel method for redefining language ecology and endangerment in NigeriaLanguage ecology, language endangerment, and relict languages: Case studies from Adamawa (Cameroon-Nigeria)A novel method for redefining language ecology and endangerment in Nigeria – towards a geospatial solutionHigh-resolution WW3 Dataset for South Western NigeriaDataSet for STEM Education in South western Nigeria

A multimodal dataset for automating language vitality and endangerment assessment in south-south Nigeria

Abstract In this paper, a multimodal dataset was collected between July 2023 and April 2

A novel method for redefining language ecology and endangerment in Nigeria

Being a multilingual and multicultural nation, Nigeria is blessed with over 525 languages (Blench, 2014) from four different language families. The sheer number of indigenous languages makes an interesting tapestry! Unfortunately, not much attention has been paid t

Language ecology, language endangerment, and relict languages: Case studies from Adamawa (Cameroon-Nigeria)

Abstract As a contribution to the more general discussion on causes of language endangerment and de

A novel method for redefining language ecology and endangerment in Nigeria – towards a geospatial solution

Being a multilingual and multicultural nation, Nigeria is blessed with over 525 languages (Blench, 2

High-resolution WW3 Dataset for South Western Nigeria

This dataset is the WW3 model hourly output of bulk wave parameters for a wave simulation for South

DataSet for STEM Education in South western Nigeria

The study investigated the relationship between teachers’ characteristics and the integration of STE