Loading...
Loading...
Speech recognition, text-to-speech, speaker identification, music classification, audio event detection
2,579 datasets
An audio dataset focused on extracting natural sounds. The description indicates the data was processed to remove music and mechanical noise. The dataset's author, organization, and specific scale are unknown.
YodaLingua-Danish is a speech dataset containing 7,871 audio-transcription pairs totaling 21 hours of Danish speech. It was created by Thomcles and is part of the multilingual YodaLingua collection. The dataset was last updated on Hugging Face in April 2026.