Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Approximately 10,000 synthetic Japanese conversation records generated for instruction tuning of large language models. Author Aratako created the dataset by applying the Magpie method to the NVIDIA Nemotron-4-340B-Instruct model via DeepInfra and published the generation code on July 5, 2024.
License is unknown; users should verify terms of use before application.