U4R/SurveyBench

Dataset

[[ Paper 📓 ]](https://arxiv.org/abs/2503.04629) [[ SurveyBench Benchmark 🤗 ]](https://huggingface.co/datasets/U4R/SurveyBench)

8

8 commits

2 linked in READMEs

updated Mar 11, 2025

See the code

README

SurveyForge: On the Outline Heuristics, Memory-Driven Generation, and Multi-dimensional Evaluation for Automated Survey Writing

[ Paper 📓 ] [ SurveyBench Benchmark 🤗 ]

🤩 Tired of chaotic structures and inaccurate references in AI-generated survey paper? SurveyForge is here to revolutionize your research experience!

We offer SurveyBench, a benchmark for academic research and evaluating the quality of AI-generated surveys.

Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:

Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:

Topics# Reference
Multimodal Large Language Models912
Evaluation of Large Language Models714
3D Object Detection in Autonomous Driving441
Vision Transformers563
Hallucination in Large Language Models500
Generative Diffusion Models994
3D Gaussian Splatting330
LLM-based Multi-Agent823
Graph Neural Networks670
Retrieval-Augmented Generation for Large Language Models608

More support topics coming soon!

🧑‍💻You can evaluate the survey from github

Contributors

yxc97

6 commits

BoZhang

2 commits

U4R/SurveyBench

Dataset

[[ Paper 📓 ]](https://arxiv.org/abs/2503.04629) [[ SurveyBench Benchmark 🤗 ]](https://huggingface.co/datasets/U4R/SurveyBench)

8

8 commits

2 linked in READMEs

updated Mar 11, 2025

See the code

README

SurveyForge: On the Outline Heuristics, Memory-Driven Generation, and Multi-dimensional Evaluation for Automated Survey Writing

[ Paper 📓 ] [ SurveyBench Benchmark 🤗 ]

🤩 Tired of chaotic structures and inaccurate references in AI-generated survey paper? SurveyForge is here to revolutionize your research experience!

We offer SurveyBench, a benchmark for academic research and evaluating the quality of AI-generated surveys.

Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:

Currently , SurveyBench consists of approximately 100 human-written survey papers across 10 distinct topics, carefully curated by doctoral-level researchers to ensure thematic consistency and academic rigor. The supported topics and the core references corresponding to each topic are as follows:

Topics# Reference
Multimodal Large Language Models912
Evaluation of Large Language Models714
3D Object Detection in Autonomous Driving441
Vision Transformers563
Hallucination in Large Language Models500
Generative Diffusion Models994
3D Gaussian Splatting330
LLM-based Multi-Agent823
Graph Neural Networks670
Retrieval-Augmented Generation for Large Language Models608

More support topics coming soon!

🧑‍💻You can evaluate the survey from github

Contributors

yxc97

6 commits

BoZhang

2 commits