Loading...
Loading...
Speech recognition, text-to-speech, speaker identification, music classification, audio event detection
2,572 datasets
NASA's 1988 Betts dataset provides site-averaged atmospheric and soil flux measurements from the FIFE field campaign. It contains data aggregated in 30-minute intervals for the single year of 1988. The dataset was collected by multiple principal investigators to study land-atmosphere interactions.
Site Averaged Neutron Soil Moisture Data: 1988 (Betts) contains daily site-averaged neutron probe soil moisture measurements from the 1987-1989 FIFE field campaign. The dataset includes only measurements from the 1988 season. NASA is the authoritative organization responsible for this data.
U-Pb zircon dating from the Canning Basin reveals a 1.7-million-year age conflict between tuffs in non-marine and marginal-marine facies, challenging established spore-pollen zonation for the middle Permian. This dataset likely contains geochemical and stratigraphic data from chemical abrasion-isotope dilution thermal ionisation mass spectrometry (CA-IDTIMS) and argon-argon dating. The findings suggest facies-specific palynofloral influences complicate correlations within the Roadian–Wordian stages.
A randomized control trial investigates the impact of music training on phonological awareness and reading skills in children with developmental dyslexia. The data supports the article by Flaugnacco, Lopez, Terribili, Montico, Zoia, and Schön. It is available via the paperswithcode platform under an Open Access license.
Scientific Data Curation Team provides metadata for a dataset of neural and physiological recordings. The data descriptor likely contains human-readable and machine-readable metadata files. The dataset's specific scale, such as participant count or recording duration, is not detailed in the provided metadata record.
A figshare-hosted dataset from a study by Lili Ming, last updated in May 2026. It contains event-related potential (ERP) data from two experiments investigating how emotional music primes affect the processing of concrete and abstract words. The dataset is small, at 26.7 KB, and is stored in an XLSX file.
A 3,315-hour collection of processed Tamil podcast audio recordings, part of a larger multilingual corpus of 57,568 hours across 12 languages. Created by InfoBayAI and last updated in June 2026, the dataset is designed to support the development of speech and conversational AI systems. It captures real-world interactions across diverse topics and formats.
Five participants completed a visual memory task under three auditory conditions during fMRI. Behavioral recall accuracy, subjective focus ratings, and region-of-interest brain activity in the anterior insula and temporo-occipital cortex were descriptively examined. The exploratory pilot study, authored by Yoshiko Tojo, suggests a paradigm for investigating affective music and memory.
125.8 MB of code and data replicates figures for a 2023 atmospheric study in Pittsburgh. Darren Cheng published this dataset under a CC-BY-4.0 license on figshare. The data includes high time resolution modeled and measured new particle formation rates.
A small dataset of 5.5 KB in XLS format, containing median levels of CS and CF before and after a participatory live music practice. Created by Nina M. van den Berg and last updated on June 1, 2026, it is shared under a CC-BY-4.0 license on figshare.
A phoneme-level manual annotation dataset for singing techniques, covering pitch and timbral dimensions. The dataset was produced as part of the SinTechSVS project by the NUS Sound and Music Computing Lab. It is associated with a paper published in IEEE/ACM TASLP 2024.
4,840 hours of processed Punjabi podcast audio form part of a larger 57,568-hour multilingual collection. The dataset captures real-world interactions across diverse topics and formats, designed to support speech and conversational AI systems. It was created by InfoBayAI and last updated on HuggingFace in June 2026.
A large-scale collection of 6,024 hours of processed Arabic podcast audio recordings, containing 57,568 hours of processed podcast audio recordings across 12 languages. It was created by InfoBayAI and last updated on 2026-06-08. The dataset captures real-world interactions across diverse topics and formats.
ASR Leaderboard Longform provides three standardized benchmark test sets—Earnings-21, Earnings-22, and TED-LIUM—for evaluating longform automatic speech recognition models. The dataset is hosted by hf-audio on Hugging Face and was last updated on June 11, 2026. It is formatted as Parquet files for efficient loading via the Hugging Face datasets library.
A large-scale collection of 2,471 hours of processed Gujarati podcast audio recordings, part of a broader multilingual corpus of 57,568 hours across 12 languages. The dataset was created by InfoBayAI and last updated in June 2026. It captures real-world interactions across diverse topics and formats to support speech AI development.
A cleaned and organized Parquet version of the bengali-tts-combined dataset, structured by speaker folders. The dataset includes audio chunks, their durations, and corresponding Bengali transcriptions, sourced from multiple speakers and videos. It was created by author 'smam' and last updated on 2026-06-15.
World Bank Group data on urban development for St. Kitts and Nevis. The dataset likely contains indicators on urbanization, traffic, congestion, and air pollution, sourced from the United Nations Population Division, World Health Organization, and other international bodies. It was last updated on 2026-04-28 and is available under a CC-BY-4.0 license.
World Bank Group data compiled from international sources like the International Road Federation and the International Telecommunications Union. This dataset likely contains indicators on water, sanitation, energy, housing, and transport infrastructure for the country of St. Kitts and Nevis. The data was last updated on 2026-04-28 and is shared under a CC-BY-4.0 license.
World Bank Group data on health systems, disease prevention, and population dynamics for St. Kitts and Nevis. The dataset covers topics including immunization, sanitation, safe drinking water, reproductive health, and nutrition. Data are aggregated from sources like the United Nations Population Division, WHO, UNICEF, and UNAIDS.
World Bank data on natural and man-made environmental resources for St. Kitts and Nevis. The dataset likely contains indicators covering forests, biodiversity, emissions, and pollution, sourced from the World Bank's data portal. It was last updated on 2026-04-28.