Google WAXAL: Automatic Speech Recognition Dataset
Available on 1 platform
Sign in to view source links and access this dataset
Description
Google WAXAL ASR Dataset is a collection of audio data for automatic speech recognition tasks. It was published on Kaggle, but its specific size, creation date, and detailed content are not provided in the available metadata. The dataset's author, organization, and license information are unknown.
Use Cases
Training acoustic models for speech-to-text systems (inferred from domain, verify after download)
Benchmarking ASR performance across different languages or accents (inferred from domain, verify after download)
Fine-tuning pre-trained speech models on specific audio characteristics (inferred from domain, verify after download)
Strengths
Published on Kaggle, a major platform for data science resources.
Limitations
Metadata is minimal; actual content requires verification after download.
Row count, file formats, and column definitions are unknown, which may limit suitability assessment.
Data may reflect geographic or source bias inherent to its collection method.
Provenance
Source
Google
License is unknown; users must verify terms before commercial use.