Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Amharic LLaMA and LLaVA: Multimodal LLMs for Low Resource Languages

Domaine:

natural language processing

Type de record:

papermodeldataset
Créateur:
And
Hôte:avatar
Large Language Models (LLMs) like GPT-4 and LLaMA have shown incredible proficiency at natural language processing tasks and have even begun to excel at tasks across other modalities such as vision and audio. Despite their success, LLMs often struggle to perform well on low-resource languages because there is so little training data available. This shortcoming is especially prevalent with open source models. In this work, we explore training LLaMA-2 to speak Amharic, a language which is spoken by over 50 million people world wide, but has orders of magnitude less data available than languages like English. We employ methods previously used for training LLMs on other languages with data scarcity, and use open source translation models to perform data augmentation and grow our dataset from millions of tokens to billions. We further enhance the capabilities of our model by connecting an image encoder and training on a translated visual instruction tuning dataset in the same manner as LLaVA, resulting in a multimodal Amharic LLM that can understand images along with text. We introduce an Amharic version of a popular benchmarking dataset to evaluate our work. Our models and dataset are open sourced and available on GitHub.

Visit

arxiv.org

Tasks

computer visionimage-text retrievallanguage modeling

Languages

Amharic

Tags

Computation and Language

Similaires

iocuydi/amharic-llama-llavaMultilingual and Multimodal LLMs in the Wild: Building for Low-Resource LanguagesBiruk-Abere/Reproducing-Amharic-LLaMA-LLaVA-PaperLLM Probe: Evaluating LLMs for Low-Resource Languageschinmayjainnnn/LLMs-for-Translation-of-Low-Resource-LanguagesLarge Multimodal Models for Low-Resource Languages: A Survey

iocuydi/amharic-llama-llava

# amharic-llama-llava Pretraining, finetuning, and inference for Amharic LLaMA and LLaVA adapted fr

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages

Multimodal LLMs are evolving from vision-language to tri-modality that see, hear, and read, yet pipe

Biruk-Abere/Reproducing-Amharic-LLaMA-LLaVA-Paper

Reproducing Amharic LLaMA and LLaVA: Multimodal LLMs for Low Resource **Abstract:-** Large

LLM Probe: Evaluating LLMs for Low-Resource Languages

Despite rapid advances in large language models (LLMs), their linguistic abilities in low-resource a

chinmayjainnnn/LLMs-for-Translation-of-Low-Resource-Languages

Machine translation from assamese to english and vice versa using state of the art LLM's # Hindi-En

Large Multimodal Models for Low-Resource Languages: A Survey

In this survey, we systematically analyze techniques used to adapt large multimodal models (LMMs) fo