Research Paper Analysis and LLM Evaluation

17 repos

Academic research repositories focused on evaluating, detecting, and analyzing large language model (LLM) capabilities and outputs, with emphasis on empirical studies and robustness testing. The cluster contains substantial research papers (TeX/BibTeX) alongside Python implementations for benchmarking, zero-shot detection methods, and systematic evaluation frameworks. Central projects include automated research agents that conduct scientific studies and tools for detecting LLM-generated content, reflecting a broader interest in understanding LLM behavior through rigorous empirical methodology.

Python · 8
TeX · 7
BibTeX Style · 1
Jupyter Notebook · 1

BigDaMa/raha

No description

TeX

70

177 commits