This is a CodeLlama-7B-Instruct-GPTQ starter template from Banana.dev that allows on-demand serverless GPU inference.
You can fork this repository and deploy it on Banana as is, or customize it based on your own needs.
For more info, check out the Banana.dev docs.
Python
82.3%
Dockerfile
17.7%
This is a CodeLlama-7B-Instruct-GPTQ starter template from Banana.dev that allows on-demand serverless GPU inference.
You can fork this repository and deploy it on Banana as is, or customize it based on your own needs.
For more info, check out the Banana.dev docs.
Python
82.3%
Dockerfile
17.7%