Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
30,000+ hours of transcribed English speech form one of the world's largest open speech recognition corpora. The dataset was created by Srijan-Chakraborty and is licensed for academic and commercial use under CC-BY-SA and CC-BY 4.0. It was last updated on HuggingFace on 2026-01-06.
License is permissive (CC-BY-SA and CC-BY 4.0), but the CC-BY-SA license requires derivative works to be shared under the same terms.