Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Targeted Lexical Injection: Unlocking Latent Cross-Lingual Alignment in Lugha-Llama via Early-Layer LoRA Fine-Tuning

Domaine:

natural language processing

Type de record:

papermodel
Créateur:
Ngu
Hôte:avatar
Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low-resource languages (LRLs), such as Swahili, often lags due to data scarcity and underrepresentation in pre-training. A key challenge is achieving robust cross-lingual lexical alignment, crucial for tasks like translation and cross-lingual information retrieval. This paper introduces Targeted Lexical Injection (TLI), a novel and efficient fine-tuning approach. We first demonstrate that Lugha-Llama-8B-wura, a Swahili-centric LLM, exhibits strong, near-perfect lexical alignment for Swahili-English word pairs in its early internal layers (specifically Layer 2, with ~0.99998 average cosine similarity based on a pilot study), a capability not fully reflected in its final output representations (baseline ~0.32 similarity on our evaluation set). TLI leverages this insight by using Low-Rank Adaptation (LoRA) and a contrastive learning objective to fine-tune the model, specifically targeting embeddings from this empirically identified optimal early layer. Our experiments show that TLI significantly improves the output-level lexical alignment for 623 trained Swahili-English word pairs, increasing average cosine similarity from 0.3211 to 0.4113 (+28.08%, p < 1.33 x 10^-240). More importantly, these improvements generalize remarkably well to 63 unseen control word pairs, with similarity increasing from 0.3143 to 0.4033 (+28.32%, p < 7.17 x 10^-27). These findings suggest TLI enhances the model's ability to preserve and propagate its inherent early-layer cross-lingual knowledge, offering a parameter-efficient and effective strategy for improving lexical alignment in LRL-focused LLMs. 11 pages, 3 figures, 2 tables. Research on parameter-efficient fine-tuning (PEFT) for low-resource languages (Swahili). Investigates cross-lingual lexical alignment in Lugha-Llama using LoRA and contrastive learning

Visit

arxiv.org

Tasks

embeddingstransfer learning

Languages

Swahili

Tags

Computation and Language68T50I.2.7; I.2.6

Similaires

Targeted Lexical Injection vs Full Fine-Tuning in Cross-Lingual Alignment for Lugha-LlamaEarly-Layer LoRA Fine-Tuning with Targeted Lexical Injection for Robustness in Low-Resource Lugha-LlamaEarly-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama and Cross-Lingual Retrieval Accuracy on Swahili BenchmarksEarly-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama: Zero-Shot Cross-Lingual Transfer Accuracy on XNLI forEarly-Layer LoRA Fine-Tuning Depth in TLI and Cross-Lingual Lexical Alignment via LAS ScoresCross-Lingual Alignment via Targeted Lexical Injection in Lugha-Llama for Zero-Shot Swahili Reasoning

Targeted Lexical Injection vs Full Fine-Tuning in Cross-Lingual Alignment for Lugha-Llama

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning with Targeted Lexical Injection for Robustness in Low-Resource Lugha-Llama

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama and Cross-Lingual Retrieval Accuracy on Swahili Benchmarks

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama: Zero-Shot Cross-Lingual Transfer Accuracy on XNLI for

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning Depth in TLI and Cross-Lingual Lexical Alignment via LAS Scores

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Cross-Lingual Alignment via Targeted Lexical Injection in Lugha-Llama for Zero-Shot Swahili Reasoning

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low