36 repos
tomg-group-umd/LoRI-D_safety_llama3_rank_32
Model Card for LoRI-D_safety_llama3_rank_32
0
6 commits
Madhurprash/Devstral-Small-2-24B-Instruct-2512-SFT-LoRA-OpenThoughts
Model Card for devstral-sft
3 commits
RampPublic/portal-mistral-7b
PorTAL refit for Mistral 7B v0.3
RampPublic/portal-qwen3-4b
PorTAL for Qwen3-4B
RampPublic/portal-qwen3-1.7b
PorTAL for Qwen3-1.7B
RampPublic/portal-qwen3-8b
PorTAL refit for Qwen3-8B
RampPublic/portal-inkling
PorTAL refit for Inkling
1
RampPublic/portal-gemma-3-4b
PorTAL cross-family refit for Gemma 3 4B
RampPublic/portal-gemma-4-e2b
PorTAL refit for Gemma 4 E2B
mzhaoshuai/Llama-2-7b-hf-conf-sft
RefAlign: RL with Similarity-based Rewards
11 commits
tuggspeedman-ai/SmolLM3-3B-summarize-dpo-lora
SmolLM3-3B-summarize-dpo-lora
tuggspeedman-ai/SmolLM3-3B-summarize-sft-lora
SmolLM3-3B-summarize-sft-lora
4 commits
zhang-manyi/ProjAtlas-xLAM2-8B-FunctionCalling-lora
ProjAtlas-xLAM2-8B-FunctionCalling-lora
solanaclawd/clawd-solana-masterpiece-qwen15-lora
Solana Clawd LoRA
huggingface/peft
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
21,691
1,857 commits
mzhaoshuai/Mistral-7B-v0.1-conf-refalign
tuggspeedman-ai/SmolVLM2-2.2B-chartqa-lora
SmolVLM2-2.2B-chartqa-lora
ArjunShukla/PromptForge-Optimizer
PromptForge-Optimizer
2
khudgins/Ornith-1.0-35B-ThinkingCap
Ornith-1.0-35B — Thinking-Cap
khudgins/Ornith-1.0-9B-ThinkingCap
Ornith-1.0-9B — Thinking-Cap
5 commits
solanaclawd/solana-nvidia-trading-factory-8b-lora
Solana NVIDIA Trading Factory LoRA
solanaclawd/solana-clawd-core-ai-1.5b-lora
Solana Clawd Core AI 1.5B LoRA
harsha-gouru/peft-hybrid-paper
Research paper: LoRA placement determines continual learning outcomes in hybrid SSM-attention…
2 commits
Shannon-Labs/shannon-control-unit
Shannon Control Unit: Adaptive regularization via control theory for LLM training
6
1 commits
paritok/paritok-4b-v1
Compresses individual long files to as little as 5% of their original size (typical content…
3
17 commits
evilfreelancer/ruGPT-3.5-13B-lora
ruGPT-3.5 13B LoRA: Adapter-Only Version
12
12 commits
hiyouga/LlamaFactory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
74,837
3,065 commits
hiyouga/LLaMA-Factory
3,070 commits
tomg-group-umd/LoRI-D_code_llama3_rank_32
Model Card for LoRI-D_code_llama3_rank_32
wuwangzhang1216/abliterix
Automated alignment adjustment for LLMs — direct steering, LoRA, and MoE expert-granular…
238
143 commits
modelscope/swift
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3,…
15,659
3,634 commits
modelscope/ms-swift
3,653 commits
jackaduma/Vicuna-LoRA-RLHF-PyTorch
A full pipeline to finetune Vicuna LLM with LoRA and RLHF on consumer hardware. Implementation of…
220
21 commits
whitecircle/halo
Halo is an open-source framework built by White Circle for training large language and multimodal…
22
16 commits
AutoArk/TinyEngram
Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.
1,374
42 commits
MakazhanAlpamys/Soup
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
6,759
961 commits