hazyresearch/lolcats-llama-3.1-8b-distill

Model

This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.

14

8 commits

1 linked in READMEs

updated Nov 24, 2024

See the code

README

This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.

Details on this model and how to train your own are provided at: https://github.com/HazyResearch/lolcats/tree/lolcats-scaled

Demo

Here is a quick GitHub GIST that will help you run inference on the model checkpoints.

Paper

See the paper page: https://huggingface.co/papers/2410.10254

text-generation

Contributors

simarora

6 commits

ariG23498

1 commits

nielsr

1 commits

hazyresearch/lolcats-llama-3.1-8b-distill

Model

This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.

14

8 commits

1 linked in READMEs

updated Nov 24, 2024

See the code

README

This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.

Details on this model and how to train your own are provided at: https://github.com/HazyResearch/lolcats/tree/lolcats-scaled

Demo

Here is a quick GitHub GIST that will help you run inference on the model checkpoints.

Paper

See the paper page: https://huggingface.co/papers/2410.10254

text-generation

Contributors

simarora

6 commits

ariG23498

1 commits

nielsr

1 commits