Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
TuniSpeech-21h is a 21-hour speech corpus designed for Tunisian Arabic (Derja). It was developed by TuniSpeech-AI to address the underrepresentation of this dialect in Automatic Speech Recognition (ASR). The dataset is compiled from social media and broadcast materials, capturing spontaneous speech and diverse linguistic characteristics.
License is unknown; terms of use must be verified before application.