Yoruba BibleTTS Aligned: Text and Audio for Speech Synthesis
Available on 1 platform
Sign in to view source links and access this dataset
Description
Yoruba BibleTTS Aligned is a dataset for text-to-speech research, likely containing aligned audio recordings and corresponding text passages. It is published on Kaggle, but the specific author, size, and creation details are not provided in the metadata. The title suggests the content is based on biblical text in the Yoruba language, a major language of Nigeria.
Use Cases
Training a text-to-speech model for Yoruba (inferred from domain, verify after download)
Developing forced alignment algorithms for audio-text pairs (inferred from domain, verify after download)
Creating pronunciation dictionaries or linguistic resources for Yoruba (inferred from domain, verify after download)
Strengths
Published on Kaggle, a major platform for sharing datasets.
Limitations
Metadata is minimal; actual content requires verification after download.
Column-level documentation is absent; field semantics must be inferred after download.
Row count, file formats, and license are unknown, which may limit suitability assessment.
Provenance
Geography
Likely related to Yoruba language speakers, primarily in Nigeria and neighboring West African regions.
License is unknown; users must verify terms of use before applying the data.