14
stars
7
commits
2
linked in READMEs
Aug 28, 2024
updated

Project Web: https://magpie-align.github.io/
Arxiv Technical Report: https://arxiv.org/abs/2406.08464
Codes: https://github.com/magpie-align/magpie
This dataset is generated by Llama 3.1 70B Instruct using Magpie. Please refer to our paper and codebase for implementation details.
This is the raw data. Feel free to apply your own filter!
License: Please follow Meta Llama 3.1 Community License.
| Model Name | Dataset | Type | Description |
|---|---|---|---|
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-1M | SFT | 1M Raw conversations built with Meta Llama 3.1 70B. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-300K-Filtered | SFT | Apply a filter and select 300K high quality conversations. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-500K-Filtered | SFT | Apply a filter and select 500K high quality conversations. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-MT-500K | SFT | Extend Magpie-Llama-3.1-Pro-500K-Filtered to multi-turn. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-MT-300K-Filtered | SFT | Select 300K high quality multi-turn conversations from Magpie-Llama-3.1-Pro-MT-500K. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-DPO-100K | DPO | DPO dataset via Best-of-N sampling and rewards. |
7 commits
14
stars
7
commits
2
linked in READMEs
Aug 28, 2024
updated

Project Web: https://magpie-align.github.io/
Arxiv Technical Report: https://arxiv.org/abs/2406.08464
Codes: https://github.com/magpie-align/magpie
This dataset is generated by Llama 3.1 70B Instruct using Magpie. Please refer to our paper and codebase for implementation details.
This is the raw data. Feel free to apply your own filter!
License: Please follow Meta Llama 3.1 Community License.
| Model Name | Dataset | Type | Description |
|---|---|---|---|
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-1M | SFT | 1M Raw conversations built with Meta Llama 3.1 70B. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-300K-Filtered | SFT | Apply a filter and select 300K high quality conversations. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-500K-Filtered | SFT | Apply a filter and select 500K high quality conversations. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-MT-500K | SFT | Extend Magpie-Llama-3.1-Pro-500K-Filtered to multi-turn. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-MT-300K-Filtered | SFT | Select 300K high quality multi-turn conversations from Magpie-Llama-3.1-Pro-MT-500K. |
| Llama 3.1 70B Instruct | Magpie-Llama-3.1-Pro-DPO-100K | DPO | DPO dataset via Best-of-N sampling and rewards. |
7 commits