6 repos
dfurman/Qwen2-72B-Orpo-v0.1
4
25 commits
dfurman/Llama-3-70B-Orpo-v0.1
2
17 commits
dfurman/CalmeRys-78B-Orpo-v0.1
79
mlx-community/CalmeRys-78B-Orpo-v0.1-4bit
0
2 commits
shibing624/DPO-En-Zh-20k-Preference
This dataset is composed by
18
5 commits
daniel-furman/sft-demos
🤖 language modeling sft demos
284 commits