Yoruba BibleTTS Aligned is a dataset for text-to-speech synthesis, likely containing aligned text and audio segments. The dataset is published on Kaggle, but detailed metadata such as author, size, and creation date are unknown. Its content appears to be derived from biblical text in the Yoruba language.
Use Cases
- Training a text-to-speech model for the Yoruba language (inferred from domain, verify after download)
- Creating a pronunciation dictionary or grapheme-to-phoneme model for Yoruba (inferred from domain, verify after download)
- Benchmarking speech synthesis alignment algorithms (inferred from domain, verify after download)
Strengths
- Published on the Kaggle platform, which provides a standard interface for access.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Column-level documentation is absent; field semantics must be inferred after download.
- Row count, file formats, and license are unknown, which may limit suitability assessment.
Provenance
- Source
- Kaggle
- Geography
- Likely focuses on the Yoruba language, which is spoken primarily in Nigeria and parts of Benin and Togo.