Loading...
Loading...
Rlhf Reward Datasets are collections for training reward models in reinforcement learning from human feedback. The dataset was authored by yitingxie and last updated on Hugging Face in January 2023. Specific details on size, format, and content are not provided in the available metadata.
License is unknown; users must verify permissions before use.