A Kaggle dataset titled 'Yoruba BibleTTS Aligned - 1TH'. The title suggests it contains Yoruba language text and corresponding audio, likely aligned for text-to-speech model training. The dataset's author, size, and specific contents are unknown from the provided metadata.
Use Cases
- Training a text-to-speech model for Yoruba (inferred from domain, verify after download)
- Evaluating forced alignment algorithms on Yoruba audio (inferred from domain, verify after download)
- Creating a Yoruba speech corpus for linguistic research (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform for sharing machine learning datasets.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Row count, file formats, and column definitions are unknown, which limits suitability assessment.
- Data may reflect source bias inherent to the specific text and recording method used.