Sign in to view source links and access this dataset
Description
A curated collection of 28,564 non-speech vocal burst audio samples. The dataset spans 18 categories, including laughter, crying, cough, and sigh. It was created by TTS-AGI and last updated on Hugging Face in March 2026.
Use Cases
Train audio classification models based on the 18 distinct vocal burst categories.
Develop emotion recognition systems based on non-verbal sounds like laughter, crying, and sighs.
Create data augmentation pipelines for speech/audio datasets based on the collection of breath, cough, and throat clearing sounds.
Benchmark generative audio models on the task of synthesizing realistic human non-speech vocalizations.
Strengths
Contains 28,564 audio samples, providing a substantial base for model training.
Covers 18 distinct categories, including common sounds like laughter (4,797 samples) and cough (4,248 samples).
Audio is provided in the FLAC format, which offers lossless compression.
Limitations
Column-level documentation is absent; field semantics must be inferred after download.
Row count is unknown, which may limit suitability assessment.
Last updated 2026-03-28 09:53:20; freshness should be verified.
Provenance
Source
TTS-AGI via Hugging Face.
Collection Method
Curated collection; specific gathering method is not detailed.
Freshness
Last updated 2026-03-28 09:53:20.
License information is unknown and should be verified before use.