This repo maintains the list of papers for repo-level code generation.
Feel free to create pull request to add more.
12/2024: FullStack Bench: Evaluating LLMs as Full Stack Coders
10/2024: RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems
10/2024: EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
6/2024: REPOEXEC: Evaluate Code Generation with a Repository-Level Executable Benchmark
2023, SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
2023, RepoEval: Repocoder: Repository-level code completion through iterative retrieval and generation
This repo maintains the list of papers for repo-level code generation.
Feel free to create pull request to add more.
12/2024: FullStack Bench: Evaluating LLMs as Full Stack Coders
10/2024: RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems
10/2024: EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
6/2024: REPOEXEC: Evaluate Code Generation with a Repository-Level Executable Benchmark
2023, SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
2023, RepoEval: Repocoder: Repository-level code completion through iterative retrieval and generation