Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
260,162 audio-transcription pairs totaling 730 hours of speech data from 13,290 distinct speakers. This Arabic portion of the YodaLingua collection is designed for training text-to-speech and automatic speech recognition models. The dataset was created by Thomcles and was last updated on May 12, 2026.
License is unknown, which may restrict commercial or research use.