Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
CraneAILabs provides a cleaned version of Luganda speech recordings from Google's WaxalNLP dataset, preprocessed for fine-tuning text-to-speech models. The dataset applies Silero VAD to remove click and pop artifacts from the start and end of audio clips, which are described as degrading model quality. This cleaned subset was last updated on March 15, 2026.
Description metadata is limited; actual data quality and structure require manual inspection after download.