6,653
stars
0
commits
37
repos using this model
20
linked in READMEs
Sep 27, 2024
updated
meta-llama/Meta-Llama-3-8B-Instruct
4,958
meta-llama/Meta-Llama-3-70B-Instruct
1,525
facebook/Meta-SecAlign-8B
14
facebook/Meta-SecAlign-70B
9
meta-llama/Meta-Llama-3.1-8B-Instruct
6,900
meta-llama/Llama-2-7b
4,537
meta-llama/Llama-2-70b
537
meta-llama/Llama-3.1-405B
987
magpie-align/magpie
[ICLR 2025] Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing. Your…
880
aryopg/mmlu-redux
32
jackyrwj/local-llm-website
Sumandora/remove-refusals-with-transformers
Implements harmful/harmless refusal removal using pure HF Transformers
2,184
ContextualAI/HALOs
A library with extensible implementations of DPO, KTO, PPO, ORPO, and other human-aware loss…
909
Vision-CAIR/MiniGPT4-video
Official code for Goldfish model for long video understanding and MiniGPT4-video for short video…
639
uclaml/SPPO
The official implementation of Self-Play Preference Optimization (SPPO)
589
likenneth/honest_llama
Inference-Time Intervention: Eliciting Truthful Answers from a Language Model
583
RLHFlow/Online-RLHF
A recipe for online RLHF and online iterative DPO.
544
neuralmagic/AutoFP8
210
Laz4rz/GPT-2
Following Karpathy with GPT-2 implementation and training, writing lots of comments cause I have…
172
ZitongYang/Synthetic_Continued_Pretraining
Code implementation of synthetic continued pretraining
162
amanb2000/Magic_Words
Code for the paper "What's the Magic Word? A Control Theory of LLM Prompting"
115
thesephist/calamity
Self-hosted GPT playground
114
amirzandieh/QJL
QJL: 1-Bit Quantized JL transform for KV Cache Quantization with Zero Overhead
100
nikolamilosevic86/local-genAI-search
Local-GenAI-Search is a generative search engine based on Llama 3, langchain and qdrant that…
94
Yu-xm/Modality_Gap_Theory
Modality Gap Theory
77
Mamba413/AdaDetectGPT
[NeurIPS 2025] AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical Guarantees
69
psunlpgroup/GreaterPrompt
GreaterPrompt: A Python Toolkit for Prompt Optimization
56
chengez/Adversarial-Paraphrasing
[NeurIPS 2025] Implementation for paper "Adversarial Paraphrasing: A Universal Attack for…
48
zjunlp/EasyEdit
[ACL 2024] An Easy-to-use Knowledge Editing Framework for LLMs.
2,915
ymcui/Chinese-LLaMA-Alpaca-3
中文羊驼大模型三期项目 (Chinese Llama-3 LLMs) developed from Meta Llama 3
1,980
Tongjilibo/bert4torch
An elegent pytorch implement of transformers
1,326
UnderstandLingBV/LLaMa2lang
Convenience scripts to finetune (chat-)LLaMa3 and other models for any language
309
SensAI-PT/LLaMa2lang
seanzhang-zhichen/llama3-chinese
Llama3-Chinese是以Meta-Llama-3-8B为底座,使用 DORA + LORA+ 的训练方法,在50w高质量中文多轮SFT数据 + 10w英文多轮SFT数据 +…
299
microsoft/FlexCAD
103
togethercomputer/Dragonfly
81
daniel-furman/sft-demos
🤖 language modeling sft demos
79
huggingface/trl-jobs
Train LLM on Hugging Face infra
72
git-disl/Virus
This is the official code for the paper "Virus: Harmful Fine-tuning Attack for Large Language…
Uminosachi/open-llm-webui
This repository contains a web application designed to execute relatively compact, locally-operated…
47
TrustMedia-zju/Lastde_Detector
Training-free LLM-generated Text Detection by Mining Token Probability Sequences (ICLR 2025)
dmis-lab/OLAPH
OLAPH: Improving Factuality in Biomedical Long-form Question Answering
PKU-Alignment/ProgressGym
Alignment with a millennium of moral progress. Spotlight@NeurIPS 2024 Track on Datasets and…
26
cambridge-mlg/jolt
16
nicolas-richet/feature-vs-text-compound-emotion
ArmelRandy/topxgen
[EMNLP 2025 Findings] TopXGen: Topic-Diverse Parallel Data Generation for Low-Resource Machine…
7
g29times/llama3-chinese
lejelly/model_merging