African Languages Exploit LLM Safety Gaps, Research Reveals Multilingual Jailbreaking Vulnerabilities

A recent study has uncovered significant vulnerabilities in the safety mechanisms of leading Large Language Models (LLMs) when subjected to multi-turn conversational attacks using low-resource African languages. Researchers specifically investigated Afrikaans, Kiswahili, isiXhosa, and isiZulu, finding that these languages can effectively bypass the guardrails of commercial LLMs such as ChatGPT, Claude, DeepSeek, Gemini, and Grok.
The methodology involved translating existing jailbreaking prompts and evaluating LLM responses through both automated testing and human red-teaming by native speakers. While simple, single-turn translation attacks proved largely ineffective, the multi-turn conversational approach yielded high success rates. For instance, harmful response rates in English ranged from 52.7% to 83.6%, in Afrikaans from 60.0% to 78.2%, and in Kiswahili from 41.8% to 70.9% across the tested models.
Crucially, the study highlighted the role of human red-teaming, which consistently increased jailbreak rates compared to automated methods. Across all evaluated languages, the average jailbreak rate surged from 59.8% to 75.8%. The research also identified translation quality as a critical factor, noting that poor translation limited the success of jailbreak attempts, indicating that improvements in translation accuracy could further amplify these vulnerabilities.
These findings carry significant implications for the development and deployment of AI in Africa. They demonstrate that LLM safety mechanisms are not robust across diverse linguistic contexts, potentially exposing African users to harmful content or misuse. The research underscores an urgent need for AI developers to enhance multilingual safety features, particularly for low-resource languages, to ensure equitable and secure AI experiences across the continent and prevent the weaponization of linguistic diversity against AI safety.
More in research
New AI Diagnostic Tool Aletheia Offers Offline Support for African Healthcare
Aletheia is an offline-first AI clinical decision support system specifically designed for low-resource healthcare settings across sub-Saharan Africa, addressing the critical lack…
New AfriSwitch Benchmark Reveals Major Gaps in AI Speech Recognition for African Code-Switched Languages
AfriSwitch is a new 61.36-hour benchmark dataset of human-transcribed, real-world code-switched speech across 16 African languages. It reveals that current AI speech recognition…
New TranslatePsy-AfriSLM Models Dramatically Improve African Language Translation for Low-Resource AI
TranslatePsy-AfriSLM introduces open-source machine translation resources for 19 Sub-Saharan African languages, including curated and synthetic data, and fine-tuned SLMs. These…
New AI Model Improves Poverty Mapping in Africa by Quantifying Uncertainty
A new machine learning method uses satellite imagery to predict poverty levels across Africa, providing crucial uncertainty estimates for policymakers. This innovation helps…
The dispatch
One email a day. The AI stories shaping Africa.
Rewritten for clarity, sourced always. No spam; unsubscribe anytime.