6 repos
lyogavin/airllm
AirLLM 70B inference with single 4GB GPU
33,779
340 commits
ManuelSLemos/RabbitLLM
Run 70B+ LLMs on a single 4GB GPU — no quantization required.
78
3 commits
HajibagheriLabs/RocketLLM
High-performance LLM engine for running massive models on consumer GPUs via asynchronous layer…
5
35 commits
gauravbatule/AirCELA
No description
3
9 commits
Mega4alik/ollm
2,798
90 commits
gmongaras/Wizard_QLoRA_Finetuning
Finetuning Some Wizard Models With QLoRA
7
14 commits