New AI System Generates Natural-Sounding Yoruba Speech

Researchers have developed TTSYoruba, a rule-based concatenative diphone speech synthesizer specifically designed for the Yoruba language. This system is integrated into the YorubaName.com online dictionary, providing audio pronunciations for personal names. It processes tone-marked Yoruba text and generates speech using a meticulously crafted phonological rule system applied to an inventory of 651 diphone units, encompassing various tonal nuances of consonant-vowel combinations.
The system's architecture addresses critical linguistic challenges unique to Yoruba, such as intricate tonal file selection logic, the disambiguation of nasal sounds (oral /n/, nasalized vowels, and syllabic nasals), and the accurate derivation of contextual rising and falling tones from level-tone inputs. This detailed phonological approach is crucial for generating highly intelligible and natural-sounding speech in a tone-rich language like Yoruba.
Beyond the technical speech synthesis, the project also contributes to Yoruba orthography by formally adopting the caron and circumflex as standard single-vowel contour tone markers. These symbols, already recognized in Yoruba phonological transcription, are now integrated into the system's text normalization pipeline and the WriteYoruba keyboard tool, enhancing the consistency and ease of writing tone-marked Yoruba. A listener study involving 50 participants evaluated the system's performance, with positive Mean Opinion Scores indicating its effectiveness.
This development is highly significant for Africa, particularly for communities that speak low-resource languages. By creating a robust text-to-speech system for Yoruba, it not only preserves and promotes the language but also opens doors for various AI applications, including educational tools, accessibility features, and localized digital content, thereby bridging the digital language divide and fostering inclusive technological advancement across the continent.
More in tools
Viamo Launches Offline AI Voice Platform in Ghana to Bridge Digital Divide
Viamo has launched the 231 Voice Platform in Ghana, offering offline AI access via toll-free phone calls to millions without smartphones or internet. This initiative uses basic…
Tether AI Unveils Offline Translation Models for 19 African Languages, Benchmarked by Peer Review
Tether AI Research has released open-source, offline translation models for 19 African languages, designed to run on basic smartphones without internet access. This addresses…
New Offline AI Diagnostic Tool Promises to Transform Healthcare in Rural Sub-Saharan Africa
Aletheia is an offline-first AI clinical decision support system specifically designed for low-resource healthcare settings in sub-Saharan Africa. It was fine-tuned on a dataset…
Intron's Sahara v2.5 Enhances African Language AI with Mid-Sentence Code-Switching and Igbo/Hausa Voice Generation
Nigerian voice AI company Intron has released Sahara v2.5, which now supports mid-sentence language switching across around 20 African languages and offers voice generation in…
The dispatch
One email a day. The AI stories shaping Africa.
Rewritten for clarity, sourced always. No spam; unsubscribe anytime.


