Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

KhmerST: A Low-Resource Khmer Scene Text Detection and Recognition Benchmark

Domaine:

natural language processing

Type de record:

paperdatasetmodel
Créateur:
NomBakLuqCou
Hôte:avatar
Developing effective scene text detection and recognition models hinges on extensive training data, which can be both laborious and costly to obtain, especially for low-resourced languages. Conventional methods tailored for Latin characters often falter with non-Latin scripts due to challenges like character stacking, diacritics, and variable character widths without clear word boundaries. In this paper, we introduce the first Khmer scene-text dataset, featuring 1,544 expert-annotated images, including 997 indoor and 547 outdoor scenes. This diverse dataset includes flat text, raised text, poorly illuminated text, distant and partially obscured text. Annotations provide line-level text and polygonal bounding box coordinates for each scene. The benchmark includes baseline models for scene-text detection and recognition tasks, providing a robust starting point for future research endeavors. The KhmerST dataset is publicly accessible at gitlab.com. Accepted at ACCV 2024

Visit

arxiv.org

Tasks

computer visionoptical character recognition

Tags

Computer Vision and Pattern Recognition

Similaires

Comprehensive Benchmark Datasets for Amharic Scene Text Detection and Recognitiongizachewteshome/Amharic-Scene-Text-Detection-and-RecognitionTowards Universal Khmer Text RecognitionThe First Swahili Language Scene Text Detection and Recognition DatasetDesalegnTIGP/Deep-learning-project-Amharic-Scene-Text-Detection-and-RecognitionAn End-to-End Scene Text Recognition for Bilingual Text

Comprehensive Benchmark Datasets for Amharic Scene Text Detection and Recognition

Ethiopic/Amharic script is one of the oldest African writing systems, which serves at least 23 langu

gizachewteshome/Amharic-Scene-Text-Detection-and-Recognition

For this project we used the only open source Amharic Scene Text Detection and Recognition dataset:

Towards Universal Khmer Text Recognition

Khmer is a low-resource language characterized by a complex script, presenting significant challenge

The First Swahili Language Scene Text Detection and Recognition Dataset

Scene text recognition is essential in many applications, including automated translation, informati

DesalegnTIGP/Deep-learning-project-Amharic-Scene-Text-Detection-and-Recognition

# Deep-learning-project-Amharic-Scene-Text-Detection-and-Recognition

An End-to-End Scene Text Recognition for Bilingual Text

Text localization and recognition from natural scene images has gained a lot of attention recently d