This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.
14
8 commits
1 linked in READMEs
updated Nov 24, 2024
This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.
Details on this model and how to train your own are provided at: https://github.com/HazyResearch/lolcats/tree/lolcats-scaled
Here is a quick GitHub GIST that will help you run inference on the model checkpoints.
See the paper page: https://huggingface.co/papers/2410.10254
This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.
14
8 commits
1 linked in READMEs
updated Nov 24, 2024
This is a pure sub-quadtratic linear attention 8B parameter model, linearized from the Meta Llama 3.1 8B model.
Details on this model and how to train your own are provided at: https://github.com/HazyResearch/lolcats/tree/lolcats-scaled
Here is a quick GitHub GIST that will help you run inference on the model checkpoints.
See the paper page: https://huggingface.co/papers/2410.10254