Large Language Model Fine-Tuning & Optimization

16 repos

Deep learning techniques for adapting and improving large language models through methods like QLoRA, DPO alignment, RLHF, and RAG pipelines. These repositories provide implementations, tutorials, and frameworks for fine-tuning LLMs efficiently on consumer hardware, optimizing model behavior through reinforcement learning and preference-based training, and augmenting models with retrieval systems. Covers both theoretical foundations in attention mechanisms and practical applied examples using PyTorch.

Jupyter Notebook · 1
llama ·3,344
machine-learning ·3,344
attention-mechanism ·3,344
language-model ·3,344
llm ·3,344
from-scratch ·3,344
gpt ·3,344
deep-learning ·3,344
educational ·3,344
natural-language-processing ·3,344