Sign in to view source links and access this dataset
Description
YouMe ASR Vosk is a dataset for automatic speech recognition (ASR) tasks, likely containing audio samples and transcriptions. It is hosted on Kaggle, but the specific content, size, and creation details are not provided in the metadata. The dataset's purpose is inferred from its title and platform.
Use Cases
Fine-tune a Vosk-based speech recognition model on new audio data (inferred from domain, verify after download)
Benchmark ASR model performance against a known corpus (inferred from domain, verify after download)
Develop audio preprocessing pipelines for speech-to-text applications (inferred from domain, verify after download)
Strengths
Published on Kaggle, a platform with established data sharing infrastructure.
Limitations
Metadata is minimal; actual content requires verification after download.
Row count, file formats, and column definitions are unknown, which limits suitability assessment.
License and authorship information are unknown, affecting terms of use.
License restrictions are unknown; verify terms before use.