Sweeping learning rate for the Pion Libera Object experiment, using their own experiment code as the base.
0
stars
2
commits
Python
primary language
Jun 16, 2026
updated
2 commits
FarshadAmiri/Fine-Tuning-on-LlaMa2
3
NVIDIA-NeMo/RL
Scalable toolkit for efficient model reinforcement
2,005
lm-playpen/playpen-paper-2025
Code for reproducing the results reported in (Horst et al., 2025)
pan-webis-de/pan-code
Code used for evaluation and baselines in the PAN shared tasks.
47
vimarsh/repro-on-the-theory-of-continual-learning-with-gradient-descent-for-neural-networks
Mimou-AJ/leo-mistral-hessianai-7b-fine-tuning-
EmilRyd/diffuse_rl
peacewang017/constrained-discrete-diffusion
Implementation of "Constrained Discrete Diffusion" (NeurIPS '2025)
86.5%
Jupyter Notebook
7.9%
Shell
5.2%