Loading...
Loading...
Speech recognition, text-to-speech, speaker identification, music classification, audio event detection
2,588 datasets
A boundary line defines the landward limit of the Ocean Management Planning Area for Massachusetts, established 0.3 nautical miles from the mean high water shoreline. The data set, provided by SCIOPS, includes official coordinate values and GIS shapefiles for this legal boundary. It was created pursuant to 'An Act Relative to Oceans' to regulate coastal development.
Records from the Massachusetts Shellfish Sanitation Program managed by the Division of Marine Fisheries (MarineFisheries). It details regulatory activities for commercial shellfish harvesting, aquaculture, and local technical partnerships. The dataset originates from the SCIOPS organization via NASA Earthdata.
Kaggle hosts this audio dataset derived from the LibriSpeech corpus. The title suggests it contains speech recordings with added noise, intended for training or evaluating MetricGAN+ and Automatic Speech Recognition systems. The dataset's author, organization, and specific details like size and license are unknown.
A dataset titled 'SLP301_MusicNet' published on Kaggle. The title suggests it contains music audio data, likely for machine learning tasks. The dataset's specific size, creator, and temporal coverage are unknown.
A three-dimensional numerical model simulates circulation in Massachusetts and Cape Cod Bays, driven by tides, wind, river runoff, and thermal forcing. The U.S. Geological Survey developed this model to study the transport of nutrients, contaminants, and red tide populations. The dataset was last updated in 1992.
Taiwan is the geographic focus of this dataset. It contains time-series data related to table tennis swings, as indicated by its title and raw description. The dataset is hosted on Kaggle, but specific details about its size, structure, and creation are currently unknown.
Project Euphonia, a public initiative led by Google, aims to improve Automatic Speech Recognition for individuals with atypical speech. The Vaani corpus expands this work beyond English to include languages such as French, Spanish, Japanese, and Hindi. This dataset is hosted by ARTPARK-IISc and was last updated on March 18, 2026.
F5-TTS Clean Voice Dataset is a collection of audio data published on Kaggle. The dataset likely contains voice recordings intended for text-to-speech model training. Its specific size, source, and creation date are not detailed in the available metadata.
Road Centerlines is a geospatial dataset representing the centerline of roadways for the City of Bloomington, Indiana, extended to a countywide network. The data includes public roads, named private roads, major multi-use trails, and proposed roadways, with attributes updated from multiple local government sources. The dataset was last updated on March 8, 2026.
Hindi Podcast Asr Dataset is a large-scale collection of raw Hindi podcast audio designed for speech and language model development. It captures real-world interactions across diverse topics and formats. The dataset was created by InfoBayAI and was last updated in March 2026.
A collection of unscripted human monologues in English, spoken by a female voice. The dataset provides 3-minute preview clips intended for use in automatic speech recognition and voice activity detection tasks. The source, author, and specific collection details are not provided.
A dataset titled 'New_music' published on the Kaggle platform. The dataset's specific content, size, and origin are not detailed in the provided metadata. Further details about the data's creator, collection method, and temporal scope require verification after accessing the dataset files.
A digital geologic-GIS dataset for the Moccasin Quadrangle in Colorado, adapted from a 1999 National Park Service geologic map. The dataset is composed of GIS data layers and tables, available in multiple formats including a file geodatabase and an OGC geopackage. It was completed as part of the NPS Geologic Resources Inventory program and includes ancillary documents with geologic unit descriptions.
A National Park Service digital geologic-GIS map for the Mancos Quadrangle in Colorado, adapted from a 1999 source map. The dataset includes GIS data layers, tables, and ancillary documents like unit descriptions and metadata. Data locational accuracy is specified to be within 12.2 meters horizontally, based on the source map scale of 1:24,000.
A National Park Service Geologic Resources Inventory digital map of the Trail Canyon Quadrangle in Colorado, adapted from a 1999 geologic map by Griffitts. The dataset includes GIS data layers and tables in multiple formats, such as a file geodatabase and geopackage, along with supporting documentation. It was produced by the NPS Geologic Resources Division as part of the Inventory and Monitoring program.
A digital geologic-GIS map of the Wetherill Mesa Quadrangle in Colorado, adapted from a 1999 National Park Service geologic map. The dataset includes GIS data layers and tables available in file geodatabase and geopackage formats, along with ancillary PDF documents containing unit descriptions and metadata. It was produced by the National Park Service's Geologic Resources Inventory program.
A digital geologic-GIS map of the Point Lookout Quadrangle in Colorado, composed of GIS data layers and tables. The dataset was produced by the National Park Service's Geologic Resources Inventory program, adapted from a 1999 source map by Griffitts. It is available in multiple GIS formats including a file geodatabase and an OGC geopackage.
A National Park Service Geologic Resources Inventory digital map of the Cortez Quadrangle, Colorado, derived from a 1999 source map. The dataset includes GIS data layers, tables, and ancillary documents like unit descriptions. Based on a source map scale of 1:24,000, features have a stated horizontal locational accuracy within 12.2 meters or 40 feet.
A digital geologic-GIS dataset for Mesa Verde National Park and vicinity, Colorado, composed of GIS data layers and tables. The data were completed as a component of the National Park Service's Geologic Resources Inventory program, adapted from source maps by Griffitts (1999). It is available in multiple GIS formats including a file geodatabase, OGC geopackage, and KMZ/KML for Google Earth.
VOXCeleb is a dataset of speech and video clips featuring celebrities. It is hosted on the Kaggle platform. The specific size, collection method, and time range are not detailed in the provided metadata.