211
stars
0
commits
10
repos using this model
1
linked in READMEs
Aug 7, 2024
updated
google/gemma-2b
1,232
google/gemma-2-27b-it
573
google/gemma-2-9b
728
google/codegemma-2b
101
google/gemma-7b
3,406
google/gemma-2b-it
949
google/shieldgemma-2b
130
google/t5gemma-2b-2b-ul2
26
unslothai/unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4,…
75,967
headroomlabs-ai/headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for…
69,076
MakazhanAlpamys/Soup
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
5,494
TransformerLensOrg/TransformerLens
A library for mechanistic interpretability of GPT-style language models
3,864
vllm-project/llm-compressor
Transformers-compatible library for applying various compression algorithms to LLMs for optimized…
3,775
tsinghua-fib-lab/AutoSOTA
674
Infini-AI-Lab/UMbreLLa
LLM Inference on consumer devices
132
ahb-sjsu/turboquant-pro
Consumer-aware compression for embedding indexes and LLM KV caches — compress by the metric the…
25
harish-kamath/rqae
Residual Quantization Autoencoder, used for interpreting LLMs
14
Parry-Parry/MechIR
Mechanistic interpretability in IR
cambridge-mlg/jolt
16