New AI System Generates Natural-Sounding Yoruba Speech
Researchers have developed TTSYoruba, a rule-based concatenative diphone speech synthesizer specifically designed for the Yoruba language. This system is integrated into the YorubaName.com online dictionary, providing audio pronunciations for personal names. It processes tone-marked Yoruba text and generates speech using a meticulously crafted phonological rule system applied to an inventory of 651 diphone units, encompassing various tonal nuances of consonant-vowel combinations.
The system's architecture addresses critical linguistic challenges unique to Yoruba, such as intricate tonal file selection logic, the disambiguation of nasal sounds (oral /n/, nasalized vowels, and syllabic nasals), and the accurate derivation of contextual rising and falling tones from level-tone inputs. This detailed phonological approach is crucial for generating highly intelligible and natural-sounding speech in a tone-rich language like Yoruba.
Beyond the technical speech synthesis, the project also contributes to Yoruba orthography by formally adopting the caron and circumflex as standard single-vowel contour tone markers. These symbols, already recognized in Yoruba phonological transcription, are now integrated into the system's text normalization pipeline and the WriteYoruba keyboard tool, enhancing the consistency and ease of writing tone-marked Yoruba. A listener study involving 50 participants evaluated the system's performance, with positive Mean Opinion Scores indicating its effectiveness.
This development is highly significant for Africa, particularly for communities that speak low-resource languages. By creating a robust text-to-speech system for Yoruba, it not only preserves and promotes the language but also opens doors for various AI applications, including educational tools, accessibility features, and localized digital content, thereby bridging the digital language divide and fostering inclusive technological advancement across the continent.
More in tools
New Open-Source AI Models Boost Speech Recognition for African Languages
DONDO is an open-source initiative providing new AI-powered automatic speech recognition models specifically for twenty-seven African language varieties across Ghana, Sierra…
M-PESA Ethiopia and Gebeya Launch AI Mini App for Mobile Users
M-PESA Ethiopia has partnered with local tech firm Gebeya to launch an AI Mini App, making various AI tools accessible to Ethiopian mobile users. This initiative aims to…
US and Morocco Launch Initiative for Advanced Drone Training and AI Innovation in North Africa
The United States and Morocco are collaborating to establish a drone academy and advanced military training center in Tan-Tan by 2030. This facility will specifically train…
Kenyan Startup Fikra API Democratizes AI Access for African Developers with Localized Payments and Pricing
Kenyan startup Fikra API has launched an AI inference platform specifically designed for African developers, addressing critical barriers like high costs, USD-only pricing, and…
The dispatch
One email a day. The AI stories shaping Africa.
Rewritten for clarity, sourced always. No spam; unsubscribe anytime.