Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Generalization of Targeted Lexical Injection Robustness to Code-Switched Social Media Text in MasakhaNER

Domain:

natural language processing

Record type:

paper
Creator:
SOV
Publisher:
Zenodo
Host:avatar
Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low-resource languages (LRLs), such as Swahili, often lags due to data scarcity and underrepresentation in pre-training. A key challenge is achieving robust cross-lingual lexical alignment, crucial for tasks like translation and cross-lingual information retrieval. This paper introduces Targeted Lexical Injection (TLI), a novel and efficient fine-tuning approach. We first demonstrate that Lugha-Llama-8B-wura, a Swahili-centric LLM, exhibits strong, near-perfect lexical alignment for Swahili-English Research goal: Does the robustness gained from Targeted Lexical Injection in Lugha-Llama generalize to code-switched social media text as measured by F1 scores on the MasakhaNER dataset? Autonomous synthesis report generated by SOVEREIGN Research Kernel. Tribunal consensus score: 7.5/10. This report was generated autonomously by SOVEREIGN Research Kernel, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 7.5/10.

Visit

doi.orgzenodo.org

Tasks

code switchinginformation extractionnamed entity recognition

Languages

Swahili

Tags

robustnessgainedTargetedLexicalInjectionLugha-Llamageneralizecode-switched

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similar

Targeted Lexical Injection and Robustness in Lugha-Llama Against AfroXLMR Code-Switching PerturbationsDetecting Propaganda Techniques in Code-Switched Social Media TextTargeted Lexical Injection via LoRA for Cross-Lingual Alignment Robustness in Swahili-English Adversarial Code-SwitchingEarly-Layer LoRA Fine-Tuning with Targeted Lexical Injection for Robustness in Low-Resource Lugha-LlamaTargeted Lexical Injection vs Full-Fine-Tuning for Cross-Lingual Representation Robustness in Low-Resource LanguagesDetecting Propaganda Techniques in Code-Switched Social Media Texts

Targeted Lexical Injection and Robustness in Lugha-Llama Against AfroXLMR Code-Switching Perturbations

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Detecting Propaganda Techniques in Code-Switched Social Media Text

Propaganda is a form of communication intended to influence the opinions and the mindset of the publ

Targeted Lexical Injection via LoRA for Cross-Lingual Alignment Robustness in Swahili-English Adversarial Code-Switching

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Early-Layer LoRA Fine-Tuning with Targeted Lexical Injection for Robustness in Low-Resource Lugha-Llama

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Targeted Lexical Injection vs Full-Fine-Tuning for Cross-Lingual Representation Robustness in Low-Resource Languages

Large Language Models (LLMs) have demonstrated remarkable capabilities, yet their performance in low

Detecting Propaganda Techniques in Code-Switched Social Media Texts

Propaganda is a planned persuasive form of communication whose goal is to influence the opinions and