Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

lilovyjgrib/X-lingual_IPA_ASR

Domaine:

natural language processing

Type de record:

model
Créateur:
lil
Hôte:
Recognize phonemes in Yoruba through a model trained on phonemes in English # X-lingual_IPA_ASR This is a final project for the Automatic Speech Recogntion course offered in Summer 2025 by the University of Tuebingen. We pretrain a BiLSTM-ResNet model on English language data from TIMIT, and evaluate its performance on Yoruba. We provide measures such as cross-entropy and PMI, hoping to disentangle learning errors from transfer errors. ## English → Yoruba and linguistic generalisation Regardless of the orthography languages draw their sounds from a universal set of types. Linguists worked out how similar these prototypical sounds are. [link] To a lage extent what sound types a language uses is studied, here we draw data from PHOIBLE. This implies that, if two languages use a similar set, some skills in recognizing the sounds of one language can be transferred to the sounds of another. How well? --- ### 📁 `PPGs/`: This folder contains scripts to extract the embeddings (last layer representation) of correctly predicted phones, as well as code to obtain their dimensionally reduced projection. The script also computes the correlation between the phone representation distance and the distance cost assigned by our fwPER formula. ### 📁 `models/`: This defines the ASRModel class, and provides helper functions for training and evaluation ### 📁 `zero-shot-final/`: Contains the paper and associated references/ images in Latex format ### 📁 `dataset/` The subfolder “mfcc_extraction_script” contains the notebooks we used to obtain the log-mels of the audio data. Data is available here: - TIMIT data. Full train set for TIMIT (logmel scale) available at: drive.google.com - Original Hugging Face Dataset huggingface.co - Yoruba data: drive.google.com ### 📁 `conversion_tools/` This contains the string processing functions, ensuring the same conventions for …

Visit

github.com

Tasks

automatic speech recognitionspeech processingtransfer learning

Languages

Yoruba

Similaires

geofact-x/GeoFact-XMasakhaNER-Xasali-xX GlueZADOKELI. EFO SELA x MAWULI ADZEI x ELIKPLIM AKORLIGeneline-X/AfricaQuantLM

geofact-x/GeoFact-X

GeoFact-X is a benchmark of geography-aware multilingual reasoning, proposed in the paper, Learn Glo

MasakhaNER-X

MasakhaNER-X is an aggregation of MasakhaNER 1.0 and MasakhaNER 2.0 datasets for 20 African language

asali-x

A cleaned and curated corpus of Hausa-language news articles from 2020-2025. This dataset is designe

X Glue

XGLUE is a new benchmark dataset to evaluate the performance of cross-lingual pre-trained models with respect to cross-lingual natural language understanding and generation. The benchmark is composed of the following 11 tasks: - NER - POS Tagging (POS) - News Class

ZADOKELI. EFO SELA x MAWULI ADZEI x ELIKPLIM AKORLI

The word “zadokeli” in Ewe, means “eclipse of the sun”. During the global pandemic in 2020, 6 eclips

Geneline-X/AfricaQuantLM

AfricaQuantLM: A browser-based demo of a 1B-parameter Gemma3 model quantized to INT4, enabling fast,