Yoruba BibleTTS Aligned is a dataset for text-to-speech research, published on Kaggle. The title suggests it contains aligned text and audio data, likely derived from Biblical text in the Yoruba language. The dataset's specific size, structure, and creation details are not provided in the available metadata.
Use Cases
- Train a text-to-speech model for Yoruba (inferred from domain, verify after download)
- Develop forced alignment tools for audio and text (inferred from domain, verify after download)
- Benchmark speech synthesis quality for African languages (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform for sharing machine learning datasets.
- The title indicates the data is aligned, which is a key feature for TTS training.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Column-level documentation is absent; field semantics must be inferred after download.
- Row count and file size are unknown, which may limit suitability assessment.