1
2 commits
1 linked in READMEs
updated Jun 11, 2024
superjelly/SPA-VL-PPO_30k
0
superjelly/SPA-VL-DPO_90k
superjelly/SPA-VL-PPO_90k
kuleshov/llama-30b-3bit
relaxml/Llama-2-70b-E8PRVQ-3Bit
relaxml/Llama-2-7b-E8PRVQ-3Bit
kuleshov/llama-30b-4bit
2
relaxml/Llama-2-13b-E8PRVQ-3Bit
EchoseChen/SPA-VL-RLHF
The reinforcement learning codes for dataset SPA-VL
48