61
stars
8
commits
2
linked in READMEs
Jan 27, 2025
updated

Project Web: https://magpie-align.github.io/
Arxiv Technical Report: https://arxiv.org/abs/2406.08464
Codes: https://github.com/magpie-align/magpie
🤨 News: Take a look on our new reasoning datasets with diverse CoT styles here!
This dataset is generated by Qwen2-72B-Instruct and Llama 3 70B Instruct using Magpie. Specifically, the instructions are generated by Qwen2-72B-Instruct, and the responses are generated by Llama 3 70B Instruct. Please refer to our paper and codebase for implementation details.
The motivation for developing this dataset is to augment the reasoning capabilities of our models through the utilization of high-quality instruction-response pairs.
You can find the model SFT checkpoint fine-tuned using this dataset here.
Please follow Meta Llama 3 Community License, Tongyi Qianwen Lincense Agreement and CC BY-NC 4.0.
If you find the model, data, or code useful, please cite our paper:
@article{xu2024magpie,
title={Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing},
author={Zhangchen Xu and Fengqing Jiang and Luyao Niu and Yuntian Deng and Radha Poovendran and Yejin Choi and Bill Yuchen Lin},
year={2024},
eprint={2406.08464},
archivePrefix={arXiv},
primaryClass={cs.CL}
}
8 commits
61
stars
8
commits
2
linked in READMEs
Jan 27, 2025
updated

Project Web: https://magpie-align.github.io/
Arxiv Technical Report: https://arxiv.org/abs/2406.08464
Codes: https://github.com/magpie-align/magpie
🤨 News: Take a look on our new reasoning datasets with diverse CoT styles here!
This dataset is generated by Qwen2-72B-Instruct and Llama 3 70B Instruct using Magpie. Specifically, the instructions are generated by Qwen2-72B-Instruct, and the responses are generated by Llama 3 70B Instruct. Please refer to our paper and codebase for implementation details.
The motivation for developing this dataset is to augment the reasoning capabilities of our models through the utilization of high-quality instruction-response pairs.
You can find the model SFT checkpoint fine-tuned using this dataset here.
Please follow Meta Llama 3 Community License, Tongyi Qianwen Lincense Agreement and CC BY-NC 4.0.
If you find the model, data, or code useful, please cite our paper:
@article{xu2024magpie,
title={Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing},
author={Zhangchen Xu and Fengqing Jiang and Luyao Niu and Yuntian Deng and Radha Poovendran and Yejin Choi and Bill Yuchen Lin},
year={2024},
eprint={2406.08464},
archivePrefix={arXiv},
primaryClass={cs.CL}
}
8 commits