This is a repo for better code generation models.
0
stars
145
commits
Python
primary language
Dec 3, 2024
updated
benchmark_human_eval.py
144 commits
1 commits
nuprl/CanItEdit
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions
51
whisperzqh/FastCoder
loubnabnl/bloom-code-evaluation
Evaluation of BLOOM on the HumanEval benchmark
6
huangd1999/EffiCoder
[ICML 2025] EffiCoder: Enhancing Code Generation in Large Language Models through Efficiency-Aware…
16
xiaoqzhwhu/ThinkCoder
1
EffiBench/EffiCoder
3
vinci-grape/APO
This is the repository for the paper titled "Aligning with Human Coding Preferences for Improving…
abcdef54/research2
93.5%
Shell
6.5%