LiquidAI/LFM2.5-2.6B-GGUF

Model

336

stars

15

commits

2

linked in READMEs

Aug 24, 2026

updated

conversational
endpoints_compatible
gguf
lfm2.5
liquid
llama.cpp
text-generation
Browse cluster: Quantized LLM Model Collections

README

Liquid AI
Try LFMDocsLEAPDiscord

LFM2.5-2.6B-GGUF

LFM2.5 is a new family of hybrid models designed for on-device deployment. It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Find more details in the original model card: https://huggingface.co/LiquidAI/LFM2.5-2.6B

🏃 How to run LFM2

Example usage with llama.cpp:

llama-cli -hf LiquidAI/LFM2.5-2.6B-GGUF --conversation \
    --temp 0.1 --top-k 50 --repeat-penalty 1.1

QAD Q4_0 GGUF

The Quantization-Aware Distillation (QAD) checkpoint is available as LFM2.5-2.6B-QAD-Q4_0.gguf.

This is distinct from the post-training-quantized LFM2.5-2.6B-Q4_0.gguf; both use the GGUF Q4_0 format.

Contributors

tuliren

8 commits

iamleonie

2 commits

darianb

1 commits

LiquidAI/LFM2.5-2.6B-GGUF

Model

336

stars

15

commits

2

linked in READMEs

Aug 24, 2026

updated

conversational
endpoints_compatible
gguf
lfm2.5
liquid
llama.cpp
text-generation
Browse cluster: Quantized LLM Model Collections

README

Liquid AI
Try LFMDocsLEAPDiscord

LFM2.5-2.6B-GGUF

LFM2.5 is a new family of hybrid models designed for on-device deployment. It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Find more details in the original model card: https://huggingface.co/LiquidAI/LFM2.5-2.6B

🏃 How to run LFM2

Example usage with llama.cpp:

llama-cli -hf LiquidAI/LFM2.5-2.6B-GGUF --conversation \
    --temp 0.1 --top-k 50 --repeat-penalty 1.1

QAD Q4_0 GGUF

The Quantization-Aware Distillation (QAD) checkpoint is available as LFM2.5-2.6B-QAD-Q4_0.gguf.

This is distinct from the post-training-quantized LFM2.5-2.6B-Q4_0.gguf; both use the GGUF Q4_0 format.

Contributors

tuliren

8 commits

iamleonie

2 commits

darianb

1 commits