Sign in to view source links and access this dataset
Description
A Yoruba language dataset for text-to-speech (TTS) applications, likely containing audio recordings and corresponding text transcripts. The dataset is titled 'BibleTTS Accepted (Consolidated)', suggesting it may be a curated collection of speech data. It is hosted on Kaggle, but detailed metadata about its size, structure, and origin is unavailable.
Use Cases
Training a text-to-speech model for the Yoruba language (inferred from domain, verify after download)
Benchmarking speech synthesis systems on religious or formal text domains (inferred from domain, verify after download)
Creating Yoruba language educational or accessibility tools (inferred from domain, verify after download)
Strengths
Published on Kaggle, a platform with an established data community.
The title suggests a consolidated, 'accepted' collection, implying a curation step.
Limitations
Metadata is minimal; actual content requires verification after download.
Row count, file formats, and column definitions are unknown.
License, author, and last update information are unavailable.
Provenance
Geography
Likely contains Yoruba language content.
License restrictions are unknown; users must verify terms of use before application.