Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
The Unsloth-DPO dataset contains question and answer pairs intended for Direct Preference Optimization (DPO) training of large language models. It was created by NeuralNovel and is inspired by the orca_dpo_pairs dataset, with a specific focus on content related to Unsloth.ai. The dataset was last updated on March 4, 2024.
License is unknown; users should verify usage rights before download.