
This project evaluates the biosecurity related capabilities and safety behaviors of large language models across English, Hausa, and Swahili. By comparing model responses to the same biosecurity prompts in different languages, the study investigates whether safety safeguards, refusals, and scientific knowledge remain consistent across the different linguistic contexts. The goal is to identify potential multilingual safety gaps and contribute to more globally representative AI biosecurity evaluations.