An Apache-2.0 dataset curated by Eric Hartford and Cognitive Computations
Discord: https://discord.gg/cognitivecomputations
Our appreciation for the generous sponsors of Dolphin R1 - Without whom this dataset could not exist.
We create a 800k sample dataset similar in composition to the one used to train DeepSeek-R1 Distill models.
The purpose of this dataset is to train R1-style reasoning models.
13 commits
An Apache-2.0 dataset curated by Eric Hartford and Cognitive Computations
Discord: https://discord.gg/cognitivecomputations
Our appreciation for the generous sponsors of Dolphin R1 - Without whom this dataset could not exist.
We create a 800k sample dataset similar in composition to the one used to train DeepSeek-R1 Distill models.
The purpose of this dataset is to train R1-style reasoning models.
13 commits