Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Magpie-Align's 150,000-instruction dataset is designed for aligning large language models, specifically for improving reasoning capabilities. It contains synthetic chain-of-thought data generated using models like DeepSeek-R1 and Llama-70B. The dataset was published in a June 2024 technical report and last updated on the Hugging Face platform in January 2025.
License terms are not specified in the provided information. The dataset is intended for research into LLM alignment, and users should review the project website and technical report for intended use and limitations.