Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
UzCrawl contains web and Telegram crawl materials from nearly 1.2 million unique sources in the Uzbek language. The dataset is updated to March 2024 and was created to support research on low-resource languages. It is a large-scale text corpus assembled by the author tahrirchi.
License information is unknown; users must verify permissions. The full description and loading script are hosted externally on Hugging Face.