7 repos
qwopqwop200/GPTQ-for-LLaMa
4 bits quantization of LLaMA using GPTQ
3,069
473 commits
IST-DASLab/gptq
Code for the ICLR 2023 paper "GPTQ: Accurate Post-training Quantization of Generative Pretrained…
2,368
32 commits
Cornell-RelaxML/quip-sharp
No description
606
46 commits
facebookresearch/any4
Quantize transformers to any learned arbitrary 4-bit numeric format
59
143 commits
artidoro/qlora
QLoRA: Efficient Finetuning of Quantized LLMs
11,012
90 commits
MajkelDcember/MS_Thesis
0
1 commits
Lightning-AI/lit-llama
Implementation of the LLaMA language model based on nanoGPT. Supports flash attention, Int8 and…
6,084
237 commits