arXiv:2305.14387 · 3 repos reference this paper in their README
tatsu-lab/alpaca_eval
2,013
·
An automatic evaluator for instruction-following language models. Human-validated, high-quality,…
tatsu-lab/alpaca_farm
844
A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human…
ScalerLab/JudgeBench
132
No description