haizelabs/Awesome-LLM-Judges

⚖️ Awesome LLM Judges ⚖️

203

12 commits

updated Apr 28, 2025

See the code

README

⚖️ Awesome LLM Judges ⚖️

This repo curates recent research on LLM Judges for automated evaluation.

[!TIP] ⚖️ Check out Verdict — our in-house library for hassle-free implementations of the papers below!


📚 Table of Contents


🌱 Starter


🎭 Multi-Judge

🤔 Debate


🎯 Finetuned Models

🌀 Hallucination

🏆 Generative Reward Models


🛡️ Safety

🛑 Content Moderation

🔍 Scalable Oversight


👨‍⚖️ Judging the Judges: Meta-Evaluation

⚖️ Biases


🤖 Agents

🚧 Coming Soon -- Stay tuned!


✨ Contributing

Have a paper to add? Found a mistake? 🧐

  • Open a pull request or submit an issue! Contributions are welcome. 🙌
  • Questions? Reach out to leonard@haizelabs.com.

Contributors

leonardtang

8 commits

qw3rtman

3 commits

dipeshbabu

1 commits

haizelabs/Awesome-LLM-Judges

⚖️ Awesome LLM Judges ⚖️

203

12 commits

updated Apr 28, 2025

See the code

README

⚖️ Awesome LLM Judges ⚖️

This repo curates recent research on LLM Judges for automated evaluation.

[!TIP] ⚖️ Check out Verdict — our in-house library for hassle-free implementations of the papers below!


📚 Table of Contents


🌱 Starter


🎭 Multi-Judge

🤔 Debate


🎯 Finetuned Models

🌀 Hallucination

🏆 Generative Reward Models


🛡️ Safety

🛑 Content Moderation

🔍 Scalable Oversight


👨‍⚖️ Judging the Judges: Meta-Evaluation

⚖️ Biases


🤖 Agents

🚧 Coming Soon -- Stay tuned!


✨ Contributing

Have a paper to add? Found a mistake? 🧐

  • Open a pull request or submit an issue! Contributions are welcome. 🙌
  • Questions? Reach out to leonard@haizelabs.com.

Contributors

leonardtang

8 commits

qw3rtman

3 commits

dipeshbabu

1 commits