Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Targeted Lexical Injection: Unlocking Latent Cross-Lingual Alignment in Lugha-Llama via Early-Layer LoRA Fine-Tuning

Domain:

natural language processing

Record type:

papermodel
Creator:
Ngu
Host:avatar
Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low-resource languages (LRLs), such as Swahili, often lags due to data scarcity and underrepresentation in pre-training. A key challenge is achieving robust cross-lingual lexical alignment, crucial for tasks like translation and cross-lingual information retrieval. This paper introduces Targeted Lexical Injection (TLI), a novel and efficient fine-tuning approach. We first demonstrate that Lugha-Llama-8B-wura, a Swahili-centric LLM, exhibits strong, near-perfect lexical alignment for Swahili-English word pairs in its early internal layers (specifically Layer 2, with ~0.99998 average cosine similarity based on a pilot study), a capability not fully reflected in its final output representations (baseline ~0.32 similarity on our evaluation set). TLI leverages this insight by using Low-Rank Adaptation (LoRA) and a contrastive learning objective to fine-tune the model, specifically targeting embeddings from this empirically identified optimal early layer. Our experiments show that TLI significantly improves the output-level lexical alignment for 623 trained Swahili-English word pairs, increasing average cosine similarity from 0.3211 to 0.4113 (+28.08%, p < 1.33 x 10^-240). More importantly, these improvements generalize remarkably well to 63 unseen control word pairs, with similarity increasing from 0.3143 to 0.4033 (+28.32%, p < 7.17 x 10^-27). These findings suggest TLI enhances the model's ability to preserve and propagate its inherent early-layer cross-lingual knowledge, offering a parameter-efficient and effective strategy for improving lexical alignment in LRL-focused LLMs. 11 pages, 3 figures, 2 tables. Research on parameter-efficient fine-tuning (PEFT) for low-resource languages (Swahili). Investigates cross-lingual lexical alignment in Lugha-Llama using LoRA and contrastive learning

Visit

arxiv.org

Tasks

embeddingstransfer learning

Languages

Swahili

Tags

Computation and Language68T50I.2.7; I.2.6

Similar

Targeted Lexical Injection vs Full Fine-Tuning in Cross-Lingual Alignment for Lugha-LlamaEarly-Layer LoRA Fine-Tuning with Targeted Lexical Injection for Robustness in Low-Resource Lugha-LlamaEarly-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama and Cross-Lingual Retrieval Accuracy on Swahili BenchmarksEarly-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama: Zero-Shot Cross-Lingual Transfer Accuracy on XNLI forEarly-Layer LoRA Fine-Tuning Depth in TLI and Cross-Lingual Lexical Alignment via LAS ScoresCross-Lingual Alignment via Targeted Lexical Injection in Lugha-Llama for Zero-Shot Swahili Reasoning

Targeted Lexical Injection vs Full Fine-Tuning in Cross-Lingual Alignment for Lugha-Llama

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning with Targeted Lexical Injection for Robustness in Low-Resource Lugha-Llama

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama and Cross-Lingual Retrieval Accuracy on Swahili Benchmarks

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning for Lexical Alignment in Lugha-Llama: Zero-Shot Cross-Lingual Transfer Accuracy on XNLI for

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning Depth in TLI and Cross-Lingual Lexical Alignment via LAS Scores

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Cross-Lingual Alignment via Targeted Lexical Injection in Lugha-Llama for Zero-Shot Swahili Reasoning

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low