Yoruba BibleTTS Aligned likely contains text and audio data for the Yoruba language. The dataset appears to be aligned for text-to-speech applications, suggesting paired Bible verses and corresponding audio clips. It was published on Kaggle, but details on its size, creation date, and author are unknown.
Use Cases
- Training a speech synthesis model for the Yoruba language (inferred from domain, verify after download)
- Creating aligned speech corpora for linguistic research (inferred from domain, verify after download)
- Benchmarking multilingual TTS systems (inferred from domain, verify after download)
Limitations
- Metadata is minimal; actual content requires verification after download
- Row count is unknown, which may limit suitability assessment
- Column-level documentation is absent; field semantics must be inferred after download
Provenance
- Geography
- Yoruba language regions (inferred from title)