Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

An embedded segmental K-means model for unsupervised segmentation and clustering of speech

Domain:

natural language processing

Record type:

papersoftware
Creator:
KamLivGol
Host:avatar
Unsupervised segmentation and clustering of unlabelled speech are core problems in zero-resource speech processing. Most approaches lie at methodological extremes: some use probabilistic Bayesian models with convergence guarantees, while others opt for more efficient heuristic techniques. Despite competitive performance in previous work, the full Bayesian approach is difficult to scale to large speech corpora. We introduce an approximation to a recent Bayesian model that still has a clear objective function but improves efficiency by using hard clustering and segmentation rather than full Bayesian inference. Like its Bayesian counterpart, this embedded segmental K-means model (ES-KMeans) represents arbitrary-length word segments as fixed-dimensional acoustic word embeddings. We first compare ES-KMeans to previous approaches on common English and Xitsonga data sets (5 and 2.5 hours of speech): ES-KMeans outperforms a leading heuristic method in word segmentation, giving similar scores to the Bayesian model while being 5 times faster with fewer hyperparameters. However, its clusters are less pure than those of the other models. We then show that ES-KMeans scales to larger corpora by applying it to the 5 languages of the Zero Resource Speech Challenge 2017 (up to 45 hours), where it performs competitively compared to the challenge baseline. 8 pages, 3 figures, 3 tables; accepted to ASRU 2017

Visit

arxiv.org

Tasks

speech processing

Languages

Tsonga

Tags

Computation and LanguageMachine Learning

Similar

Application Of K-Means Clustering For Customer Segmentation In Grocery Stores In Kenyaclustering an african hairstyle dataset using pca and k-meansOkes2024/A-hybrid-AI-model-integrating-LSTM-XGBoost-and-K-means-for-interpretable-prediction-_-clusteringApplication of k Means Clustering algorithm for prediction of Students Academic PerformanceA Hybrid Machine Learning Model for Predicting Surgical Procedure Duration: Integrating Random Forest and K-Means ClusteringClustering and Classification of Cotton Lint Using Principle Component Analysis, Agglomerative Hierarchical Clustering, and K-Means Clustering

Application Of K-Means Clustering For Customer Segmentation In Grocery Stores In Kenya

The retail industry, particularly in the context of grocery stores, plays a vital role in meeting co

clustering an african hairstyle dataset using pca and k-means

The adoption of digital transformation was not expressed in building an African face shape classifie

Okes2024/A-hybrid-AI-model-integrating-LSTM-XGBoost-and-K-means-for-interpretable-prediction-_-clustering

A Hybrid Machine Learning Framework for Water Quality Assessment and Contamination Clustering in the

Application of k Means Clustering algorithm for prediction of Students Academic Performance

The ability to monitor the progress of students academic performance is a critical issue to the acad

A Hybrid Machine Learning Model for Predicting Surgical Procedure Duration: Integrating Random Forest and K-Means Clustering

International audience Efficient operating room (OR) management depends on the accura

Clustering and Classification of Cotton Lint Using Principle Component Analysis, Agglomerative Hierarchical Clustering, and K-Means Clustering

Cotton from the three cotton growing regions of Uganda was characterized for 13 quality parameters u