3
5 commits
1 linked in READMEs
updated Jun 3, 2025
lmms-lab/MME
Evaluation Dataset for MME
37
evalplus/mbppplus
19
MRMRbenchmark/design
The evaluation code is implemented based on [MTEB](https://github.com/embeddings-benchmark/mteb/tre…
0
homebrewltd/Maze-Bench-v0.2
1
MRMRbenchmark/theorem
limingcv/MultiGen-20M_canny_eval
RewardMATH/RewardMATH
evalplus/evalperf
tongye98/Awesome-Code-Benchmark
A comprehensive code domain benchmark review of LLM researches.
245