Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

A Ngemba Sociocultural Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
Ins
Host:
A Ngemba Sociocultural Dataset is a sociocultural and onomastic dataset documenting the Ngemba community of Cameroon (administratively known as bafoussamIII-Bmoungoum, self-designated Bamoungoum/mϋngum), collected under the project "Projet de confection d'un didacticiel sur la composition des peuplements au Cameroun" (a didactic-software project on the composition of Cameroon's settled peoples), coordinated by Prof. Emmanuel Ngue Um for CERDOTOLA (the Inter-State Institution of scientific co-operation for the preservation, diffusion and valorization of the African cultural heritage) during the project's pilot data-collection phase in March-April 2020. Unlike a purely linguistic corpus, the dataset foregrounds the community itself as the unit of documentation: its history, social organisation, religious life and material culture, with onomastics, common expressions and a numeral system providing a linguistic component elicited in the community's own speech (recorded in the International Phonetic Alphabet, or in a Latin-based approximation where IPA was not available to the fieldworker). The Ngemba (Bamoungpoum) are a majority community who trace their origin to Bansoa. They have an ongoing territorial dispute with the Bameka. Ancestor cults are actively maintained. The biennial festival *nə kaŋ pə muŋgum* — spanning four months in odd years — is the principal communal and ritual gathering. The founding ancestor is *fɛndzɔnveu* (the hunter). The dataset's primary added value lies in documenting the Ngemba community's own self-description — its social organisation, kinship and marriage practices, religious and symbolic life, economy and material culture, and internal system of personal, ethnic and place names — rather than only its named language or dialect. This community-centred, ethnographic design distinguishes it from language-only documentation, while the accompanying onomastic, common-expression and numeral-system data still make the dataset usable for lexicographic and comparative linguistic work on the Ngemba variety.

Visit

mozilladatacollective.com

Languages

Ngemba

Tags

mdcmozilla data collectiveOTHTSVPDF

Licenses

Nwulite Obodo Open Data Licence 1.0 (NOODL-1.0)

Similar

A Gbaya Sociocultural DatasetA Bum Sociocultural DatasetA Kanuri Sociocultural DatasetA Bakoko Sociocultural DatasetA Ghomala Sociocultural DatasetA Bulu Sociocultural Dataset

A Gbaya Sociocultural Dataset

A Gbaya Sociocultural Dataset is a sociocultural and onomastic dataset documenting the Gbaya communi

A Bum Sociocultural Dataset

A Bum Sociocultural Dataset is a sociocultural and onomastic dataset documenting the Bum (also spell

A Kanuri Sociocultural Dataset

A Kanuri Sociocultural Dataset is a sociocultural and onomastic dataset documenting the Kanuri commu

A Bakoko Sociocultural Dataset

A Bakoko Sociocultural Dataset is a sociocultural and onomastic dataset documenting the Bakoko commu

A Ghomala Sociocultural Dataset

A Ghomala Sociocultural Dataset is a sociocultural and onomastic dataset documenting the Ghomala (se

A Bulu Sociocultural Dataset

A Bulu Sociocultural Dataset is a sociocultural and onomastic dataset documenting the Bulu community