CudaLLM: A Language Model for High-Performance CUDA Kernel Generation
29
4 commits
1 linked in READMEs
updated Aug 3, 2025
cudaLLM-8B is a language model for generating high-performance and syntactically correct CUDA kernels. It is based on the Qwen3-8B model and has undergone a two-stage training process to master the complexities of parallel programming for GPUs.
Performance on KernelBench:
| Bo1 | Bo2 | Bo4 | Bo8 | Bo16 | |
|---|---|---|---|---|---|
| Level-1 | 79.75 | 83 | 84 | 86 | 87 |
| Level-2 | 67.30 | 70 | 71 | 72 | 73 |
| Level-3 | 20.83 | 26 | 30 | 34 | 36 |
The model was trained using the verl library. The model was trained and evaluated on:
The primary use of CudaLLM is to assist developers in writing and optimizing high-performance CUDA kernels. It can be used for:
4 commits
CudaLLM: A Language Model for High-Performance CUDA Kernel Generation
29
4 commits
1 linked in READMEs
updated Aug 3, 2025
cudaLLM-8B is a language model for generating high-performance and syntactically correct CUDA kernels. It is based on the Qwen3-8B model and has undergone a two-stage training process to master the complexities of parallel programming for GPUs.
Performance on KernelBench:
| Bo1 | Bo2 | Bo4 | Bo8 | Bo16 | |
|---|---|---|---|---|---|
| Level-1 | 79.75 | 83 | 84 | 86 | 87 |
| Level-2 | 67.30 | 70 | 71 | 72 | 73 |
| Level-3 | 20.83 | 26 | 30 | 34 | 36 |
The model was trained using the verl library. The model was trained and evaluated on:
The primary use of CudaLLM is to assist developers in writing and optimizing high-performance CUDA kernels. It can be used for:
4 commits