Aligned text and audio data for Yoruba Bible text-to-speech synthesis, published on Kaggle. The dataset likely contains paired text transcripts and corresponding audio recordings. Specifics on the number of samples, data collection method, and contributors are not provided in the metadata.
Use Cases
- Training a speech synthesis model for Yoruba (inferred from domain, verify after download)
- Creating a pronunciation dictionary or grapheme-to-phoneme aligner (inferred from domain, verify after download)
- Fine-tuning a multilingual TTS system (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform for sharing machine learning datasets.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Row count, file formats, and column definitions are unknown, which may limit suitability assessment.