A dataset likely containing aligned text and audio for the Yoruba language, intended for text-to-speech (TTS) applications. It is hosted on the Kaggle platform. The specific source, size, and creation details are not provided in the available metadata.
Use Cases
- Train a text-to-speech model for Yoruba (inferred from domain, verify after download)
- Study phoneme or prosody alignment in Yoruba speech (inferred from domain, verify after download)
- Benchmark speech synthesis models on religious or formal text domains (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform for sharing data science resources.
- The title suggests the data is aligned, which is a key requirement for training TTS models.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Column-level documentation is absent; field semantics must be inferred after download.
- Row count, file formats, and license are unknown, which may limit suitability assessment.