Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
616 hours of English audio extracted from the Emilia-Dataset, licensed under CC BY 4.0. The audio events are classified using Scribe v1, an STT/ASR system from ElevenLabs, and filtered using Facebook audio aesthetics metrics. The dataset is described as a 'v1' version, with further collaboration invited via Discord.
License for the core dataset is CC BY 4.0, but full transaction timestamps from Scribe v1 are noted as CC BY 4.0 NC (non-commercial).