Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
A distilled dataset from Deepseek-R1 based on medical verifiable problems, released on 2025/02/22. The dataset is designed for instruction tuning and Supervised Fine-Tuning (SFT) of language models. It was created by author vBasura and last updated on 2026-05-23.
The dataset page notes a split: 'medical_o1_sft.json' contains only medical SFT data, while 'medical_o1_sft_mix.json' contains a mix of medical and general instruction data.