Yoruba BibleTTS Aligned is a dataset for text-to-speech research, likely containing aligned audio recordings and corresponding text transcripts. It was published on Kaggle, but the author, organization, and specific data volume are unknown. The dataset's last update date is also unspecified.
Use Cases
- Train a text-to-speech model for the Yoruba language (inferred from domain, verify after download)
- Develop forced alignment tools for audio and text data (inferred from domain, verify after download)
- Benchmark speech synthesis quality for religious or formal text domains (inferred from domain, verify after download)
Strengths
- Published on Kaggle, a platform with a large community for data sharing.
- Title indicates the dataset provides aligned text and audio, a key feature for TTS model training.
Limitations
- Metadata is minimal; actual content requires verification after download.
- Row count, file formats, and column definitions are unknown, which may limit suitability assessment.
- Data may reflect bias inherent to its specific source material (e.g., religious text).