Logo Lanfrica

UniversalDependencies/UD_Hausa-EasternAutogramm

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Uni
Hôte:
# Summary This treebank contains data of the Autogramm project, for the (Kano) Eastern dialect of Hausa, Nigeria. # Introduction These samples of Eastern Hausa are transcripts of Hausa news broadcasts of the BBC World Service. The treebank is maintained in the SUD framework: SUD_Hausa-EasternAutogramm and converted automatically in UD. The treebank contains 18 samples, 335 trees, 9,820 tokens and 9,032 words. # Acknowledgments The samples are extracts from (Jaggar 1992), a Hausa textbook published by SOAS. The translations and annotations are by B. Caron. ## References Jaggar, Philip J. 1992. An advanced Hausa reader with grammatical notes and exercises. London: School of Oriental and African Studies, University of London. # Changelog * 2026-05-15 v2.18 * Initial release in Universal Dependencies. === Machine-readable metadata (DO NOT REMOVE!) ================================ Data available since: UD v2.18 License: CC BY-SA 4.0 Includes text: yes Parallel: no Genre: news Lemmas: manual native UPOS: manual native XPOS: not available Features: manual native Relations: manual native Contributors: Caron, Bernard Contributing: elsewhere Contact: bernard.l.caron@gmail.com ===============================================================================

Languages