Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Averroes-Q: Training a 32-Billion Parameter Bilingual Arabic-English LLM on a Single Apple M2 Ultra Workstation

Domaine:

natural language processing

Type de record:

model
Créateur:
Yah
Éditeur:
Zenodo
Hôte:avatar
We present Averroes-Q, a 32-billion parameter bilingual Arabic-English large language model fine-tuned entirely on a single Apple M2 Ultra workstation with 192 GB of unified memory. Built upon the Qwen2.5-32B-Instruct architecture, Averroes-Q is the first publicly released Arabic-capable LLM trained exclusively on consumer-grade Apple Silicon hardware. Using parameter-efficient fine-tuning via Low-Rank Adaptation (LoRA) and Apple's MLX framework, we demonstrate that a complete bilingual instruction-tuning pipeline—including training, adapter fusion, quantization, and deployment—can be executed on a single machine without any cloud GPU infrastructure. The model is trained on the Averroes Corpus, a curated bilingual dataset of 5.7 million instruction-response pairs, with a filtered high-quality v2 subset of 1.2 million examples. Over four training iterations spanning adapter ranks 8 and 16 with up to 5,000 steps, we achieve a best validation loss of 1.440 and a training loss of 1.629. Peak memory consumption reaches 67.4 GB, well within the M2 Ultra's 192 GB capacity. Averroes-Q serves as the primary bilingual backbone for multiple production systems, including the SAIF framework, the Rushd assistant, and the Hayula Bot. We release all model variants, training configurations, and evaluation benchmarks to the research community via HuggingFace, where the model has received over 156 downloads. Our results demonstrate that competitive bilingual LLMs can be developed on consumer hardware, significantly lowering the barrier to entry for low-resource language AI research.

Visit

doi.org

Tasks

language modeling

Tags

Arabic NLPAverroesmodel trainingbilingual modelsartificial intelligence

Licenses

info:eu-repo/semantics/openAccessCreative Commons Attribution Non Commercial Share Alike 4.0 Internationalhttps://creativecommons.org/licenses/by-nc-sa/4.0/legalcode

Similaires

Parameter-Efficient Fine-Tuning for LLM-Based Arabic-to-English Machine TranslationParameter setting and feature mismatch in a Yoruba-English bilingual childArabicWeb-Edu: Educational Quality Data for Arabic LLM TrainingLongitudinal predictors of single word spelling in Northern Sotho-English bilingual children: a cross-linguistic studyDeveloping a Bilingual English-Arabic Dataset for Textbook Question Answering: A Hybrid Translation and Validation ApproachGhana Maternal Health Q&A Dataset (English)

Parameter-Efficient Fine-Tuning for LLM-Based Arabic-to-English Machine Translation

Large Language Models (LLMs) such as GPT-3, BLOOM, BERT... have revolutionized natural language proc

Parameter setting and feature mismatch in a Yoruba-English bilingual child

anguage acquisition studies on bilingual children within the African context are rare. Furthermore,

ArabicWeb-Edu: Educational Quality Data for Arabic LLM Training

The quality of training data plays a critical role in the performance of large language models (LLMs

Longitudinal predictors of single word spelling in Northern Sotho-English bilingual children: a cross-linguistic study

Abstract Although there is overwhelming evidence highlighting the foundational role of p

Developing a Bilingual English-Arabic Dataset for Textbook Question Answering: A Hybrid Translation and Validation Approach

Textbook Question Answering has been a central feature of educational artificial intelligence enabli

Ghana Maternal Health Q&A Dataset (English)

20,000 English Questions & Answers for Maternal Health