[[ Paper 📓 ]](https://arxiv.org/abs/2503.04629) [[ SurveyBench Benchmark 🤗 ]](https://huggingface.co/datasets/U4R/SurveyBench)
8
8 commits
2 linked in READMEs
updated Mar 11, 2025
We offer SurveyBench, a benchmark for academic research and evaluating the quality of AI-generated surveys.
Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:
Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:
| Topics | # Reference |
|---|---|
| Multimodal Large Language Models | 912 |
| Evaluation of Large Language Models | 714 |
| 3D Object Detection in Autonomous Driving | 441 |
| Vision Transformers | 563 |
| Hallucination in Large Language Models | 500 |
| Generative Diffusion Models | 994 |
| 3D Gaussian Splatting | 330 |
| LLM-based Multi-Agent | 823 |
| Graph Neural Networks | 670 |
| Retrieval-Augmented Generation for Large Language Models | 608 |
More support topics coming soon!
[[ Paper 📓 ]](https://arxiv.org/abs/2503.04629) [[ SurveyBench Benchmark 🤗 ]](https://huggingface.co/datasets/U4R/SurveyBench)
8
8 commits
2 linked in READMEs
updated Mar 11, 2025
We offer SurveyBench, a benchmark for academic research and evaluating the quality of AI-generated surveys.
Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:
Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:
| Topics | # Reference |
|---|---|
| Multimodal Large Language Models | 912 |
| Evaluation of Large Language Models | 714 |
| 3D Object Detection in Autonomous Driving | 441 |
| Vision Transformers | 563 |
| Hallucination in Large Language Models | 500 |
| Generative Diffusion Models | 994 |
| 3D Gaussian Splatting | 330 |
| LLM-based Multi-Agent | 823 |
| Graph Neural Networks | 670 |
| Retrieval-Augmented Generation for Large Language Models | 608 |
More support topics coming soon!