Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Mequanent/amharic-tokenizer-fast

Record type:

model
Creator:
Meq
Host:

Visit

huggingface.co

Languages

Amharic

Similar

misge10/amharic-tokenizerBonnief/amharic-nllb-tokenizerdagim19/amharic-tokenizer-bpesefineh-ai/Amharic-Tokenizergashudemman/amharic-sentencepiece-tokenizerAmharic Segmenter and tokenizer

misge10/amharic-tokenizer

Bonnief/amharic-nllb-tokenizer

dagim19/amharic-tokenizer-bpe

This a code to train a bpe tokenizer using the amharic alphabets and amharic letters. The trained to

sefineh-ai/Amharic-Tokenizer

Syllable-aware BPE tokenizer for the Amharic language (አማርኛ) – fast, accurate, trainable. # Amharic

gashudemman/amharic-sentencepiece-tokenizer

Amharic Segmenter and tokenizer

This is a simple script that split an Amharic document into different sentences and tokenes. If you find an issue, please let us know in the GitHub (https://github.com/uhh-lt/amharicprocessor/issues)