Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
A dataset likely containing human preference labels for training reinforcement learning from human feedback (RLHF) models. The dataset is published on Kaggle, but specific details about its size, creation date, and authors are not provided in the available metadata. Its title suggests it is related to the 'Helpful and Harmless' (HH) benchmark for aligning language models.
License is unknown; users must verify terms of use before applying the data.