Yoruba BibleTTS Aligned is a dataset for text-to-speech research, published on Kaggle. The title suggests it contains aligned audio and text data, likely derived from biblical text in the Yoruba language. Specifics on size, author, and update date are unavailable.
Use Cases
- Train a text-to-speech model for Yoruba (inferred from domain, verify after download)
- Benchmark forced alignment algorithms on a low-resource language corpus (inferred from domain, verify after download)
- Fine-tune a multilingual speech model on Yoruba audio (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a major platform for data science.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Column-level documentation is absent; field semantics must be inferred after download.
- Row count is unknown, which may limit suitability assessment.