Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Automated tone transcription

Domain:

natural language processing

Record type:

papersoftware
Creator:
Bir
Host:avatar
In this paper I report on an investigation into the problem of assigning tones to pitch contours. The proposed model is intended to serve as a tool for phonologists working on instrumentally obtained pitch data from tone languages. Motivation and exemplification for the model is provided by data taken from my fieldwork on Bamileke Dschang (Cameroon). Following recent work by Liberman and others, I provide a parametrised F_0 prediction function P which generates F_0 values from a tone sequence, and I explore the asymptotic behaviour of downstep. Next, I observe that transcribing a sequence X of pitch (i.e. F_0) values amounts to finding a tone sequence T such that P(T) {}~= X. This is a combinatorial optimisation problem, for which two non-deterministic search techniques are provided: a genetic algorithm and a simulated annealing algorithm. Finally, two implementations---one for each technique---are described and then compared using both artificial and real data for sequences of up to 20 tones. These programs can be adapted to other tone languages by adjusting the F_0 prediction function. 12 pages, 4 postscript figures, uses examples.sty, newapa.sty, latex-acl.sty, ipamacs.sty

Visit

arxiv.org

Languages

Yemba

Tags

Computation and Language

Similar

Automated Transcription of Gə'əz Manuscripts Using Deep LearningGigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinementkenngare/swahili-transcriptionrahelFM/Transcription-LesothoChristophe4514/nlp-transcriptionSpeech transcription server

Automated Transcription of Gə'əz Manuscripts Using Deep Learning

This paper describes a collaborative project designed to meet the needs of communities interested in

GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement

The evolution of speech technology has been spurred by the rapid increase in dataset sizes. Traditio

kenngare/swahili-transcription

use llm to transcribe swahili audio # Swahili Audio Transcription A Python tool for transcribing S

rahelFM/Transcription-Lesotho

## Speech-to-Text Benchmarking on Code-Switched isiZulu-English Dataset This repository contains a n

Christophe4514/nlp-transcription

Encoder-Decoder Architecture in RNNs for language translate (french-lingala) # NLP-Transcription&Tr

Speech transcription server

This is the "Parliament-specific" application server component implemented as a proof-of-concept dur