Yoruba BibleTTS Aligned is a dataset for text-to-speech research, likely containing aligned audio recordings and corresponding text from the Bible in the Yoruba language. It was published on Kaggle, but specific details about its size, creation date, and author are unknown. The dataset's primary purpose appears to be training and evaluating speech synthesis models for Yoruba.
Use Cases
- Train a neural text-to-speech model for Yoruba (inferred from domain, verify after download)
- Benchmark speech synthesis quality on religious or formal text (inferred from domain, verify after download)
- Develop pronunciation or prosody models for Yoruba phonetics (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform with established data sharing infrastructure.
- Focuses on Yoruba, a language that may be underrepresented in speech synthesis resources.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Row count, file formats, and audio specifications are unknown, which may limit suitability assessment.
- Column-level documentation is absent; field semantics must be inferred after download.