work in progress
0
stars
10
commits
Python
primary language
Aug 7, 2026
updated
adapted from https://github.com/dllm-reasoning/d1/tree/main and https://github.com/cychomatica/FreeDave
10 commits
EmilRyd/diffuse_rl
NVIDIA-NeMo/RL
Scalable toolkit for efficient model reinforcement
2,005
hoangNguyen210/Reinforcement-Learning-CBM
To RL-CBM
ChloeL19/RLVF
Reinforcement Learning from Verifier Feedback in Coq
1
kavin525zhang/LLM_reasoning_framework
hjx620/paddlenlp
LoselSpt/LLM
zsoltii/ai
95.0%
Shell
5.0%