evanmiller/LLM-Reading-List

LLM papers I'm reading, mostly on inference and model compression

751

13 commits

updated Dec 21, 2023

See the code

README

Just helping myself keep track of LLM papers that I‘m reading, with an emphasis on inference and model compression.

Transformer Architectures

Foundation Models

Position Encoding

KV Cache

Activation

Pruning

Quantization

Normalization

Sparsity and rank compression

Fine-tuning

Sampling

Scaling

Mixture of Experts

Watermarking

More

Contributors

evanmiller

13 commits

evanmiller/LLM-Reading-List

LLM papers I'm reading, mostly on inference and model compression

751

13 commits

updated Dec 21, 2023

See the code

README

Just helping myself keep track of LLM papers that I‘m reading, with an emphasis on inference and model compression.

Transformer Architectures

Foundation Models

Position Encoding

KV Cache

Activation

Pruning

Quantization

Normalization

Sparsity and rank compression

Fine-tuning

Sampling

Scaling

Mixture of Experts

Watermarking

More

Contributors

evanmiller

13 commits