250+ Fine-tuning & RL Notebooks for text, vision, audio, embedding, TTS models.
5,665
stars
1,161
commits
Jupyter Notebook
primary language
Sep 4, 2026
updated
Below are Colab notebooks, organized by model. You can also view all notebooks in our docs.
The notebooks run locally and feature data prep, training and inference. Read our fine-tuning guide.
| Model | Type | Notebook Link |
|---|---|---|
| Qwen2.5 Coder (1.5B) | Tool Calling | |
| FunctionGemma (270M) | Tool Calling | |
| FunctionGemma (270M) | Mobile Actions | |
| FunctionGemma (270M) | Inference | |
| FunctionGemma (270M) | Conversational |
| Model | Type | Notebook Link |
|---|---|---|
| Spark TTS (0.5B) | TTS | |
| Llasa TTS (1B) | TTS | |
| Orpheus (3B) | TTS | |
| Llasa TTS (3B) | TTS | |
| Sesame CSM (1B) | TTS | |
| Oute TTS (1B) | TTS |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning | |
| Paddle OCR (1B) | Vision |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| gpt oss MXFP4 (20B) | Inference | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss BNB (20B) | Inference | |
| (A100) gpt oss (120B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| LFM2.5 (1.2B) | Conversational | |
| Liquid LFM2 (1.2B) | Conversational | |
| Liquid LFM2 | Conversational | |
| LFM2.5 VL (1.6B) | Vision | |
| LFM2.5 (1.2B) | Translation | |
| Falcon H1 (0.5B) | Alpaca | |
| Falcon H1 | Alpaca |
| Model | Type | Notebook Link |
|---|---|---|
| (A100) Nemotron Nano 3 30B A3B | Conversational | |
| (A100) Nemotron 3 Nano 30B A3B | Conversational |
| Model | Type | Notebook Link |
|---|---|---|
| Muse Glimmer (30B) | Conversational | |
| Muse Glimmer (30B) | Vision | |
| Muse Glimmer (30B) | ||
| CodeForces CoT Reasoning | ||
| Synthetic Data Hackathon | Synthetic Data | |
| Unsloth | Studio |
| Model | Type | Notebook Link |
|---|---|---|
| Spark TTS (0.5B) | TTS | |
| Llasa TTS (1B) | TTS | |
| Orpheus (3B) | TTS | |
| Llasa TTS (3B) | TTS | |
| Sesame CSM (1B) | TTS | |
| Oute TTS (1B) | TTS |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning | |
| Paddle OCR (1B) | Vision |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| gpt oss MXFP4 (20B) | Inference | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss BNB (20B) | Inference | |
| (A100) gpt oss (120B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| (A100) Nemotron Nano 3 30B A3B | Conversational | |
| (A100) Nemotron 3 Nano 30B A3B | Conversational |
These notebooks target AMD ROCm GPUs and are not available in Colab. View / download them directly from GitHub:
| Model | Type | Notebook |
|---|---|---|
| Unsloth Studio | Chat UI | GitHub |
| Gemma4 (E2B) | Vision | GitHub |
| Qwen3 5 (4B) | Vision | GitHub |
| Qwen3 5 (2B) | Vision | GitHub |
| gpt oss (20B) | Fine Tuning | GitHub |
| gpt oss (20B) | Auto Kernel Creation | GitHub |
| Model | Type | Notebook |
|---|---|---|
| Qwen3 (0 6B) | Phone Deployment | GitHub |
| Qwen3 (0.6B) | Reasoning Conversational | GitHub |
| Llama3.1 (8B) | Inference | GitHub |
| Llama3.1 (8B) | GSM8K Math + vLLM | GitHub |
| NeMo Gym Sudoku | Sudoku | GitHub |
| NeMo Gym Multi Environment | Multi Environment | GitHub |
| Whisper (Large) | Fine Tuning | GitHub |
| gpt oss MXFP4 (20B) | Inference | GitHub |
| gpt oss BNB (20B) | Inference | GitHub |
| gpt oss BF16 (20B) | 2048 Game | GitHub |
| gpt oss (20B) | 2048 Game | GitHub |
| gpt oss (20B) | Minesweeper Game | GitHub |
| gpt oss (20B) | Auto Kernel Creation | GitHub |
| gpt oss (20B) | Fine Tuning | GitHub |
| (OpenEnv) gpt oss BF16 (20B) | 2048 Game | GitHub |
| (OpenEnv) gpt oss (20B) | 2048 Game | GitHub |
| (DGX Spark) gpt oss (20B) | 2048 Game | GitHub |
| Spark TTS (0.5B) | TTS | GitHub |
| Qwen3 (8B) | DAPO Math + vLLM | GitHub |
| Llama3 (8B) | Ollama | GitHub |
| Llama3 (8B) | ORPO | GitHub |
| Llama3 (8B) | Alpaca | GitHub |
| Openenv wordle | Wordle + vLLM | GitHub |
| gpt oss (20B) | Auto Kernel Creation | GitHub |
| gpt oss (120B) | Fine Tuning | GitHub |
| Qwen2.5 (3B) | GSM8K Math + vLLM | GitHub |
| ModernBert | Classification | GitHub |
| Qwen3 (4B) | QAT | GitHub |
| Qwen3 (4B) | Conversational | GitHub |
| Qwen2.5 VL (7B) | Vision Math + vLLM | GitHub |
| Qwen2.5 VL (7B) | Vision | GitHub |
| Llasa TTS (1B) | TTS | GitHub |
| Llama3.2 (1B) | DAPO Math + vLLM | GitHub |
| Llama3.2 (1B) | RAFT | GitHub |
| Deepseek OCR (3B) | Fine Tuning | GitHub |
| Deepseek OCR (3B) | Evaluation | GitHub |
| Deepseek OCR (3B) | Eval | GitHub |
| Paddle OCR (1B) | Vision | GitHub |
| ERNIE 4 5 VL 28B A3B PT | Vision | GitHub |
| Deepseek OCR 2 (3B) | Fine Tuning | GitHub |
| Qwen3 VL (8B) | Vision | GitHub |
| Qwen3 VL (8B) | Vision Math | GitHub |
| Mistral v0.3 (7B) | GSM8K Math + vLLM | GitHub |
| Mistral v0.3 (7B) | Conversational | GitHub |
| Qwen3 5 MoE | MoE | GitHub |
| Orpheus (3B) | TTS | GitHub |
| Llasa TTS (3B) | TTS | GitHub |
| Meta Synthetic Data Llama3.1 (8B) | GRPO | GitHub |
| Meta Synthetic Data Llama3 2 (3B) | GRPO | GitHub |
| Llama3.2 (1B and 3B) | Conversational | GitHub |
| Qwen 3 5 27B(80GB) | Conversational | GitHub |
| Qwen3 (32B) | Reasoning Conversational | GitHub |
| Llama3.1 (8B) | Alpaca | GitHub |
| Qwen3 5 (4B) | Vision Math | GitHub |
| Qwen3 (14B) | Conversational | GitHub |
| Qwen3 (14B) | Reasoning Conversational | GitHub |
| CodeForces CoT Reasoning | GitHub | |
| Llama3.3 (70B) | Conversational | GitHub |
| Synthetic Data Hackathon | Synthetic Data | GitHub |
| DiffusionGemma (26B A4B) | Sudoku | GitHub |
| Gemma3 (4B) | Conversational | GitHub |
| Gemma3 (4B) | Vision Math | GitHub |
| Phi 4 (14B) | GSM8K Math + vLLM | GitHub |
| Phi 4 | Conversational | GitHub |
| Gemma3 (27B) | Conversational | GitHub |
| Qwen3 5 (0 8B) | Vision | GitHub |
| GLM Flash(80GB) | Conversational | GitHub |
| Sesame CSM (1B) | TTS | GitHub |
| Gemma4 (31B) | Conversational | GitHub |
| Gemma4 (31B) | Vision | GitHub |
| Qwen2 VL (7B) | Vision | GitHub |
| Qwen3 (4B) | Thinking | GitHub |
| Qwen3 MoE | MoE | GitHub |
| Gemma3 (1B) | GSM8K Math | GitHub |
| Nemotron Nano 3 30B A3B | Conversational | GitHub |
| Nemotron 3 Nano 30B A3B | Conversational | GitHub |
| Gemma4 (E4B) | Conversational | GitHub |
| Gemma4 (E4B) | Vision | GitHub |
| Gemma4 (E4B) | Audio | GitHub |
| Llama3.2 (11B) | Vision | GitHub |
| Phi 3.5 Mini | Conversational | GitHub |
| Gemma4 (26B A4B) | Conversational | GitHub |
| Gemma4 (26B A4B) | Vision | GitHub |
| Magistral (24B) | Reasoning Conversational | GitHub |
| Qwen2.5 (7B) | Alpaca | GitHub |
| DeepSeek R1 0528 Qwen3 (8B) | DAPO Math + vLLM | GitHub |
| Qwen2.5 Coder (1.5B) | Tool Calling | GitHub |
| Gemma3 (270M) | Conversational | GitHub |
| Gemma3 (270M) | Phone Deployment | GitHub |
| Qwen2.5 Coder (14B) | Conversational | GitHub |
| FunctionGemma (270M) | Conversational | GitHub |
| FunctionGemma (270M) | Inference | GitHub |
| FunctionGemma (270M) | Tool Calling | GitHub |
| FunctionGemma (270M) | Mobile Actions | GitHub |
| Gemma3N (4B) | Multimodal | GitHub |
| Gemma3N (4B) | Audio | GitHub |
| LFM2.5 (1.2B) | DAPO Math | GitHub |
| LFM2.5 (1.2B) | Conversational | GitHub |
| Gemma4 (E2B) | Conversational | GitHub |
| Gemma4 (E2B) | Sudoku | GitHub |
| Gemma4 (E2B) | 2048 Game | GitHub |
| Gemma4 (E2B) | Auto Kernel Creation | GitHub |
| Gemma4 (E2B) | Audio | GitHub |
| Qwen3 6 MoE | MoE | GitHub |
| Qwen3 Embedding (4B) | Embeddings | GitHub |
| Qwen3 (4B) | DAPO Math + vLLM | GitHub |
| Pixtral (12B) | Vision | GitHub |
| Qwen3 Embedding (0 6B) | Embeddings | GitHub |
| Mistral Small (22B) | Alpaca | GitHub |
| Liquid LFM2 (1.2B) | Conversational | GitHub |
| Liquid LFM2 | Conversational | GitHub |
| LFM2.5 VL (1.6B) | Vision | GitHub |
| Ministral3 VL (3B) | Vision | GitHub |
| Ministral3 (3B) | Sudoku | GitHub |
| Gemma3 (4B) | Vision | GitHub |
| Oute TTS (1B) | TTS | GitHub |
| Llama3 (8B) | Conversational | GitHub |
| ERNIE 4 5 21B A3B PT | Conversational | GitHub |
| Granite4.0 (3B) | Conversational | GitHub |
| Qwen3 (14B) | Alpaca | GitHub |
| LFM2.5 (1.2B) | Translation | GitHub |
| LFM2.5 (1.2B) | Text Completion | GitHub |
| Gemma3N (4B) | Vision | GitHub |
| Granite4.0 (350M) | Conversational | GitHub |
| TinyLlama (1.1B) | Alpaca | GitHub |
| Falcon H1 (0.5B) | Alpaca | GitHub |
| Falcon H1 | Alpaca | GitHub |
| Phi 3 Medium | Conversational | GitHub |
| EmbeddingGemma (300M) | Embeddings | GitHub |
| Gemma2 (9B) | Alpaca | GitHub |
| Gemma2 (2B) | Alpaca | GitHub |
| Mistral v0.3 (7B) | CPT | GitHub |
| Mistral v0.3 (7B) | Alpaca | GitHub |
| Mistral (7B) | Text Completion | GitHub |
| Qwen2 (7B) | Alpaca | GitHub |
| Zephyr (7B) | DPO | GitHub |
| ModernBERT (Large) | Classification | GitHub |
| Mistral Nemo (12B) | Alpaca | GitHub |
| CodeGemma (7B) | Conversational | GitHub |
| BGE M3 | Embeddings | GitHub |
| TinyQwen3 MoE | MoE | GitHub |
| All MiniLM L6 v2 | Embeddings | GitHub |
Run any of these on molab, Marimo's hosted GPU notebooks. They're reactive: change a value in one cell, the cells below recompute on their own.
| Model | Type | Notebook |
|---|---|---|
| Gemma4 (E2B) | Vision | |
| Qwen3 5 (4B) | Vision | |
| Qwen3 5 (2B) | Vision | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss (20B) | Auto Kernel Creation | |
| Qwen3 (14B) | Reasoning Conversational |
If you'd like to contribute to our notebooks, here's a guide to get you started:
Template_Notebook.ipynb in the root directory of this project. This template contains the basic structure and formatting guidelines for all notebooks in this collection.Template_Notebook.ipynb.<Model Name>-<Type>.ipynb (e.g., Mistral_v0.3_(7B)-Alpaca.ipynb)<Model Name>-Vision.ipynb (e.g., Llava_v1.6_(7B)-Vision.ipynb)<Type>: Alpaca, Conversational, CPT, DPO, ORPO, Text_Completion, CSV, Inference, Unsloth_Studiooriginal_template: Once your notebook is ready, move it to the original_template directory.python update_all_notebooks.py
This script will automatically:
original_template to the notebooks directory.README.md file.(top 30 of 32)
Jupyter Notebook
93.1%
Python
6.9%
250+ Fine-tuning & RL Notebooks for text, vision, audio, embedding, TTS models.
5,665
stars
1,161
commits
Jupyter Notebook
primary language
Sep 4, 2026
updated
Below are Colab notebooks, organized by model. You can also view all notebooks in our docs.
The notebooks run locally and feature data prep, training and inference. Read our fine-tuning guide.
| Model | Type | Notebook Link |
|---|---|---|
| Qwen2.5 Coder (1.5B) | Tool Calling | |
| FunctionGemma (270M) | Tool Calling | |
| FunctionGemma (270M) | Mobile Actions | |
| FunctionGemma (270M) | Inference | |
| FunctionGemma (270M) | Conversational |
| Model | Type | Notebook Link |
|---|---|---|
| Spark TTS (0.5B) | TTS | |
| Llasa TTS (1B) | TTS | |
| Orpheus (3B) | TTS | |
| Llasa TTS (3B) | TTS | |
| Sesame CSM (1B) | TTS | |
| Oute TTS (1B) | TTS |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning | |
| Paddle OCR (1B) | Vision |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| gpt oss MXFP4 (20B) | Inference | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss BNB (20B) | Inference | |
| (A100) gpt oss (120B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| LFM2.5 (1.2B) | Conversational | |
| Liquid LFM2 (1.2B) | Conversational | |
| Liquid LFM2 | Conversational | |
| LFM2.5 VL (1.6B) | Vision | |
| LFM2.5 (1.2B) | Translation | |
| Falcon H1 (0.5B) | Alpaca | |
| Falcon H1 | Alpaca |
| Model | Type | Notebook Link |
|---|---|---|
| (A100) Nemotron Nano 3 30B A3B | Conversational | |
| (A100) Nemotron 3 Nano 30B A3B | Conversational |
| Model | Type | Notebook Link |
|---|---|---|
| Muse Glimmer (30B) | Conversational | |
| Muse Glimmer (30B) | Vision | |
| Muse Glimmer (30B) | ||
| CodeForces CoT Reasoning | ||
| Synthetic Data Hackathon | Synthetic Data | |
| Unsloth | Studio |
| Model | Type | Notebook Link |
|---|---|---|
| Spark TTS (0.5B) | TTS | |
| Llasa TTS (1B) | TTS | |
| Orpheus (3B) | TTS | |
| Llasa TTS (3B) | TTS | |
| Sesame CSM (1B) | TTS | |
| Oute TTS (1B) | TTS |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning | |
| Paddle OCR (1B) | Vision |
| Model | Type | Notebook Link |
|---|---|---|
| Deepseek OCR (3B) | Fine Tuning | |
| Deepseek OCR (3B) | Evaluation | |
| Deepseek OCR (3B) | Eval | |
| Deepseek OCR 2 (3B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| gpt oss MXFP4 (20B) | Inference | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss BNB (20B) | Inference | |
| (A100) gpt oss (120B) | Fine Tuning |
| Model | Type | Notebook Link |
|---|---|---|
| (A100) Nemotron Nano 3 30B A3B | Conversational | |
| (A100) Nemotron 3 Nano 30B A3B | Conversational |
These notebooks target AMD ROCm GPUs and are not available in Colab. View / download them directly from GitHub:
| Model | Type | Notebook |
|---|---|---|
| Unsloth Studio | Chat UI | GitHub |
| Gemma4 (E2B) | Vision | GitHub |
| Qwen3 5 (4B) | Vision | GitHub |
| Qwen3 5 (2B) | Vision | GitHub |
| gpt oss (20B) | Fine Tuning | GitHub |
| gpt oss (20B) | Auto Kernel Creation | GitHub |
| Model | Type | Notebook |
|---|---|---|
| Qwen3 (0 6B) | Phone Deployment | GitHub |
| Qwen3 (0.6B) | Reasoning Conversational | GitHub |
| Llama3.1 (8B) | Inference | GitHub |
| Llama3.1 (8B) | GSM8K Math + vLLM | GitHub |
| NeMo Gym Sudoku | Sudoku | GitHub |
| NeMo Gym Multi Environment | Multi Environment | GitHub |
| Whisper (Large) | Fine Tuning | GitHub |
| gpt oss MXFP4 (20B) | Inference | GitHub |
| gpt oss BNB (20B) | Inference | GitHub |
| gpt oss BF16 (20B) | 2048 Game | GitHub |
| gpt oss (20B) | 2048 Game | GitHub |
| gpt oss (20B) | Minesweeper Game | GitHub |
| gpt oss (20B) | Auto Kernel Creation | GitHub |
| gpt oss (20B) | Fine Tuning | GitHub |
| (OpenEnv) gpt oss BF16 (20B) | 2048 Game | GitHub |
| (OpenEnv) gpt oss (20B) | 2048 Game | GitHub |
| (DGX Spark) gpt oss (20B) | 2048 Game | GitHub |
| Spark TTS (0.5B) | TTS | GitHub |
| Qwen3 (8B) | DAPO Math + vLLM | GitHub |
| Llama3 (8B) | Ollama | GitHub |
| Llama3 (8B) | ORPO | GitHub |
| Llama3 (8B) | Alpaca | GitHub |
| Openenv wordle | Wordle + vLLM | GitHub |
| gpt oss (20B) | Auto Kernel Creation | GitHub |
| gpt oss (120B) | Fine Tuning | GitHub |
| Qwen2.5 (3B) | GSM8K Math + vLLM | GitHub |
| ModernBert | Classification | GitHub |
| Qwen3 (4B) | QAT | GitHub |
| Qwen3 (4B) | Conversational | GitHub |
| Qwen2.5 VL (7B) | Vision Math + vLLM | GitHub |
| Qwen2.5 VL (7B) | Vision | GitHub |
| Llasa TTS (1B) | TTS | GitHub |
| Llama3.2 (1B) | DAPO Math + vLLM | GitHub |
| Llama3.2 (1B) | RAFT | GitHub |
| Deepseek OCR (3B) | Fine Tuning | GitHub |
| Deepseek OCR (3B) | Evaluation | GitHub |
| Deepseek OCR (3B) | Eval | GitHub |
| Paddle OCR (1B) | Vision | GitHub |
| ERNIE 4 5 VL 28B A3B PT | Vision | GitHub |
| Deepseek OCR 2 (3B) | Fine Tuning | GitHub |
| Qwen3 VL (8B) | Vision | GitHub |
| Qwen3 VL (8B) | Vision Math | GitHub |
| Mistral v0.3 (7B) | GSM8K Math + vLLM | GitHub |
| Mistral v0.3 (7B) | Conversational | GitHub |
| Qwen3 5 MoE | MoE | GitHub |
| Orpheus (3B) | TTS | GitHub |
| Llasa TTS (3B) | TTS | GitHub |
| Meta Synthetic Data Llama3.1 (8B) | GRPO | GitHub |
| Meta Synthetic Data Llama3 2 (3B) | GRPO | GitHub |
| Llama3.2 (1B and 3B) | Conversational | GitHub |
| Qwen 3 5 27B(80GB) | Conversational | GitHub |
| Qwen3 (32B) | Reasoning Conversational | GitHub |
| Llama3.1 (8B) | Alpaca | GitHub |
| Qwen3 5 (4B) | Vision Math | GitHub |
| Qwen3 (14B) | Conversational | GitHub |
| Qwen3 (14B) | Reasoning Conversational | GitHub |
| CodeForces CoT Reasoning | GitHub | |
| Llama3.3 (70B) | Conversational | GitHub |
| Synthetic Data Hackathon | Synthetic Data | GitHub |
| DiffusionGemma (26B A4B) | Sudoku | GitHub |
| Gemma3 (4B) | Conversational | GitHub |
| Gemma3 (4B) | Vision Math | GitHub |
| Phi 4 (14B) | GSM8K Math + vLLM | GitHub |
| Phi 4 | Conversational | GitHub |
| Gemma3 (27B) | Conversational | GitHub |
| Qwen3 5 (0 8B) | Vision | GitHub |
| GLM Flash(80GB) | Conversational | GitHub |
| Sesame CSM (1B) | TTS | GitHub |
| Gemma4 (31B) | Conversational | GitHub |
| Gemma4 (31B) | Vision | GitHub |
| Qwen2 VL (7B) | Vision | GitHub |
| Qwen3 (4B) | Thinking | GitHub |
| Qwen3 MoE | MoE | GitHub |
| Gemma3 (1B) | GSM8K Math | GitHub |
| Nemotron Nano 3 30B A3B | Conversational | GitHub |
| Nemotron 3 Nano 30B A3B | Conversational | GitHub |
| Gemma4 (E4B) | Conversational | GitHub |
| Gemma4 (E4B) | Vision | GitHub |
| Gemma4 (E4B) | Audio | GitHub |
| Llama3.2 (11B) | Vision | GitHub |
| Phi 3.5 Mini | Conversational | GitHub |
| Gemma4 (26B A4B) | Conversational | GitHub |
| Gemma4 (26B A4B) | Vision | GitHub |
| Magistral (24B) | Reasoning Conversational | GitHub |
| Qwen2.5 (7B) | Alpaca | GitHub |
| DeepSeek R1 0528 Qwen3 (8B) | DAPO Math + vLLM | GitHub |
| Qwen2.5 Coder (1.5B) | Tool Calling | GitHub |
| Gemma3 (270M) | Conversational | GitHub |
| Gemma3 (270M) | Phone Deployment | GitHub |
| Qwen2.5 Coder (14B) | Conversational | GitHub |
| FunctionGemma (270M) | Conversational | GitHub |
| FunctionGemma (270M) | Inference | GitHub |
| FunctionGemma (270M) | Tool Calling | GitHub |
| FunctionGemma (270M) | Mobile Actions | GitHub |
| Gemma3N (4B) | Multimodal | GitHub |
| Gemma3N (4B) | Audio | GitHub |
| LFM2.5 (1.2B) | DAPO Math | GitHub |
| LFM2.5 (1.2B) | Conversational | GitHub |
| Gemma4 (E2B) | Conversational | GitHub |
| Gemma4 (E2B) | Sudoku | GitHub |
| Gemma4 (E2B) | 2048 Game | GitHub |
| Gemma4 (E2B) | Auto Kernel Creation | GitHub |
| Gemma4 (E2B) | Audio | GitHub |
| Qwen3 6 MoE | MoE | GitHub |
| Qwen3 Embedding (4B) | Embeddings | GitHub |
| Qwen3 (4B) | DAPO Math + vLLM | GitHub |
| Pixtral (12B) | Vision | GitHub |
| Qwen3 Embedding (0 6B) | Embeddings | GitHub |
| Mistral Small (22B) | Alpaca | GitHub |
| Liquid LFM2 (1.2B) | Conversational | GitHub |
| Liquid LFM2 | Conversational | GitHub |
| LFM2.5 VL (1.6B) | Vision | GitHub |
| Ministral3 VL (3B) | Vision | GitHub |
| Ministral3 (3B) | Sudoku | GitHub |
| Gemma3 (4B) | Vision | GitHub |
| Oute TTS (1B) | TTS | GitHub |
| Llama3 (8B) | Conversational | GitHub |
| ERNIE 4 5 21B A3B PT | Conversational | GitHub |
| Granite4.0 (3B) | Conversational | GitHub |
| Qwen3 (14B) | Alpaca | GitHub |
| LFM2.5 (1.2B) | Translation | GitHub |
| LFM2.5 (1.2B) | Text Completion | GitHub |
| Gemma3N (4B) | Vision | GitHub |
| Granite4.0 (350M) | Conversational | GitHub |
| TinyLlama (1.1B) | Alpaca | GitHub |
| Falcon H1 (0.5B) | Alpaca | GitHub |
| Falcon H1 | Alpaca | GitHub |
| Phi 3 Medium | Conversational | GitHub |
| EmbeddingGemma (300M) | Embeddings | GitHub |
| Gemma2 (9B) | Alpaca | GitHub |
| Gemma2 (2B) | Alpaca | GitHub |
| Mistral v0.3 (7B) | CPT | GitHub |
| Mistral v0.3 (7B) | Alpaca | GitHub |
| Mistral (7B) | Text Completion | GitHub |
| Qwen2 (7B) | Alpaca | GitHub |
| Zephyr (7B) | DPO | GitHub |
| ModernBERT (Large) | Classification | GitHub |
| Mistral Nemo (12B) | Alpaca | GitHub |
| CodeGemma (7B) | Conversational | GitHub |
| BGE M3 | Embeddings | GitHub |
| TinyQwen3 MoE | MoE | GitHub |
| All MiniLM L6 v2 | Embeddings | GitHub |
Run any of these on molab, Marimo's hosted GPU notebooks. They're reactive: change a value in one cell, the cells below recompute on their own.
| Model | Type | Notebook |
|---|---|---|
| Gemma4 (E2B) | Vision | |
| Qwen3 5 (4B) | Vision | |
| Qwen3 5 (2B) | Vision | |
| gpt oss (20B) | Fine Tuning | |
| gpt oss (20B) | Auto Kernel Creation | |
| Qwen3 (14B) | Reasoning Conversational |
If you'd like to contribute to our notebooks, here's a guide to get you started:
Template_Notebook.ipynb in the root directory of this project. This template contains the basic structure and formatting guidelines for all notebooks in this collection.Template_Notebook.ipynb.<Model Name>-<Type>.ipynb (e.g., Mistral_v0.3_(7B)-Alpaca.ipynb)<Model Name>-Vision.ipynb (e.g., Llava_v1.6_(7B)-Vision.ipynb)<Type>: Alpaca, Conversational, CPT, DPO, ORPO, Text_Completion, CSV, Inference, Unsloth_Studiooriginal_template: Once your notebook is ready, move it to the original_template directory.python update_all_notebooks.py
This script will automatically:
original_template to the notebooks directory.README.md file.(top 30 of 32)
Jupyter Notebook
93.1%
Python
6.9%