Large Language Model Fine-tuning & Evaluation

10 repos

Tools and frameworks for adapting and evaluating large language models, particularly focused on instruction-following and specialized task performance. The cluster centers on open-source implementations of model fine-tuning pipelines (SCRIBE variants based on Llama models), evaluation frameworks for reasoning and long-context tasks, and educational resources for understanding model adaptation techniques. Projects here span both the practical engineering of model customization and the assessment methods needed to measure their effectiveness.

Jupyter Notebook · 1
TypeScript · 1
reasoning ·1,344
education ·932
theory-of-mind ·931
prompt-engineering ·931
hacktoberfest ·931
o1 ·931
machine-learning ·931
ai ·931
literacy ·931
tutoring ·931