4,537
stars
0
commits
32
repos using this model
10
linked in READMEs
Apr 17, 2024
updated
meta-llama/Llama-2-70b
537
meta-llama/Llama-2-7b-chat
625
meta-llama/Llama-2-13b-chat
295
meta-llama/Llama-2-7b-chat-hf
4,837
llamaste/Llama-2-7b-chat-hf
llamaste/Llama-2-70b-chat-hf
2,210
meta-llama/CodeLlama-7b-hf
128
llamaste/Llama-2-13b-chat-hf
1,124
galilai-group/llm-jepa
337
ayaka14732/llama-2-jax
JAX implementation of the Llama 2 model
217
hamelsmu/llama-inference
experiments with inference on llama
103
sehyunkwon/ICTC
This is a public repository for Image Clustering Conditioned on Text Criteria (IC|TC)
93
kubaxipl11/ml-animations
🎨 Explore interactive visualizations of Machine Learning and Linear Algebra concepts with a…
7
MikailSio/mikailsarpkayawebsite
1
shankarpandala/learn-llm
RightNow-AI/autokernel
Autoresearch for GPU kernels. Give it any PyTorch model, go to sleep, wake up to optimized Triton…
1,556
AnswerDotAI/fsdp_qlora
Training LLMs with QLoRA + FSDP
1,552
microsoft/Samba
[ICLR 2025] Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language…
972
jzhang38/EasyContext
Memory optimization and training recipes to extrapolate language models' context length to 1…
761
datamllab/LongLM
[ICML'24 Spotlight] LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning
668
jy-yuan/KIVI
[ICML 2024] KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache
432
wesg52/world-models
Extracting spatial and temporal world models from LLMs
263
princeton-nlp/HELMET
The HELMET Benchmark
227
aounon/llm-rank-optimizer
132
tsinghua-fib-lab/AAAI2025_MIA-Tuner
[AAAI'25 Oral] "MIA-Tuner: Adapting Large Language Models as Pre-training Text Detector".
130
microsoft/FlexCAD
zhuochunli/Learn-from-Committee
The code for paper "Learning from Committee: Reasoning Distillation from a Mixture of Teachers with…
pranavjad/tinyllama-bitnet
Train your own small bitnet model
86
EleutherAI/gpt-neox
An implementation of model parallel autoregressive transformers on GPUs, based on the Megatron and…
7,461
meta-pytorch/gpt-fast
Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.
6,251
turingmotors/heron
Heron is a library that seamlessly integrates multiple Vision and Language models, as well as Video…
178
ictnlp/TruthX
Code for ACL 2024 paper "TruthX: Alleviating Hallucinations by Editing Large Language Models in…
145
quic/cloud-ai-sdk
Qualcomm Cloud AI SDK (Platform and Apps) enable high performance deep learning inference on…
85
ZhuoyangLiu2005/MLA
MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic…
75
Marker-Inc-Korea/KO-Platypus
[KO-Platy🥮] Korean-Open-platypus를 활용하여 llama-2-ko를 fine-tuning한 KO-platypus model
73
cambridge-mlg/jolt
16
openlangrid/mlgrid-services
A framework to serve various generative AIs and other machine learning functions via WebSocket and…
2
milesway/scientific_llm
LARGE LANGUAGE MODELS FOR SCIENTIFIC PROBLEM SOLVING