Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
RAID is a benchmark dataset for evaluating machine-generated text detectors, containing over 10 million documents. It spans 11 large language models, 11 genres, 4 decoding strategies, and 12 adversarial attacks. The dataset was created by liamdugan and updated in September 2024.
The full description and specific data structure are hosted externally on the Hugging Face dataset page.