Yoruba BibleTTS Aligned - LEV is a dataset for text-to-speech research, published on Kaggle. The title suggests it contains aligned text and audio data, likely derived from the Yoruba translation of the Bible's Book of Leviticus. Specific details on the number of utterances, audio format, and creation methodology are not provided in the available metadata.
Use Cases
- Train a TTS model on aligned Yoruba audio and text (inferred from domain, verify after download)
- Benchmark speech synthesis quality for low-resource languages (inferred from domain, verify after download)
- Study prosody and pronunciation in Yoruba religious texts (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform for sharing machine learning datasets.
- The title indicates the data is aligned, which is a key requirement for TTS model training.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Column-level documentation is absent; field semantics must be inferred after download.
- Row count, audio specifications, and license are unknown, which may limit suitability assessment.