Yoruba BibleTTS Aligned is a dataset for speech synthesis, likely containing audio recordings aligned with corresponding text. The dataset is hosted on Kaggle, but detailed metadata such as the number of samples, file formats, and creation details are not provided. Its title suggests a focus on the Yoruba language and the Book of Exodus.
Use Cases
- Training a speech synthesis model for Yoruba (inferred from domain, verify after download)
- Developing forced alignment tools for audio-text pairs (inferred from domain, verify after download)
- Studying prosody and pronunciation in Yoruba religious texts (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform with established data sharing infrastructure.
- The title indicates a specific alignment between audio and text, which is a foundational feature for TTS datasets.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Column-level documentation is absent; field semantics must be inferred after download.
- Row count, file formats, and license are unknown, which may limit suitability assessment.