Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
A Russian-language dataset for binary classification of texts as toxic or non-toxic. It contains 72,468 training records, balanced 1:1, with 9,059 validation and 9,059 test records, created by author koryl. The dataset was last updated on 2026-07-16.
The full description and data sources are truncated; refer to the dataset page at https://huggingface.co/datasets/koryl/russian_toxicity_dataset for complete details.