roboflow/notebooks

A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3, and Qwen3-VL.

9,673

stars

655

commits

Jupyter Notebook

primary language

Aug 14, 2026

updated

roboflow.com/models
automatic-labeling-system
computer-vision
deep-learning
deep-neural-networks
google-colab
image-classification
image-segmentation
machine-learning
object-detection
open-vocabulary-detection
open-vocabulary-segmentation
paligemma
pytorch
qwen
tutorial
vlm
yolov5
yolov8
zero-shot-classification
zero-shot-detection

README

👋 hello

This repository offers a growing collection of computer vision tutorials. Learn to use SOTA models like YOLOv11, SAM 2, Florence-2, PaliGemma 2, and Qwen2.5-VL for tasks ranging from object detection, segmentation, and pose estimation to data extraction and OCR. Dive in and explore the exciting world of computer vision!

🚀 model tutorials (61 notebooks)

notebookopen in colab / kaggle / sagemaker studio labcomplementary materialsrepository / paper
RF-DETR Keypoint Detection and Fine-TuningColab KaggleGitHub
Object Detection with Google Gemini 3.5 FlashColab Kaggle
How to Perform OCR with GLM-OCRColab KaggleYouTubeGitHub arXiv
How to Track Objects with RF-DETR and ByteTrack TrackerColab KaggleYouTubeGitHub arXiv
Fine-Tune YOLO26 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLO26 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Segment Images with SAM3Colab KaggleRoboflow YouTubeGitHub arXiv
Segment Videos with SAM3Colab KaggleRoboflow YouTubeGitHub arXiv
Open Vocabulary Object Detection with Qwen3-VLColab KaggleGitHub
Fine-Tune RF-DETR Segmentation on Custom DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Zero-Shot Object Detection and Segmentation with Google Gemini 2.5Colab KaggleRoboflowarXiv
Fine-Tune RF-DETR on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Zero-Shot Object Detection and Segmentation with YOLOEColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv12 on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Zero-Shot Object Detection with Qwen2.5-VLColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune Qwen2.5-VL for JSON Data ExtractionColab KaggleYouTubeGitHub arXiv
Fine-Tune PaliGemma2 on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Fine-Tune PaliGemma2 for JSON Data ExtractionColab KaggleRoboflowGitHub arXiv
Fine-Tune PaliGemma2 for LaTeX OCRColab KaggleRoboflowGitHub arXiv
Fine-Tune SAM-2.1Colab KaggleRoboflow YouTubeGitHub
Fine-Tune GPT-4o on Object Detection DatasetColab KaggleRoboflow YouTube
Fine-Tune YOLO11 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLO11 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Segment Images with SAM2Colab KaggleRoboflow YouTubeGitHub arXiv
Segment Videos with SAM2Colab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune RT-DETR on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Fine-Tune Florence-2 on Object Detection DatasetColab KaggleRoboflow YouTubearXiv
Run Different Vision Tasks with Florence-2Colab KaggleRoboflow YouTubearXiv
Fine-Tune PaliGemma on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv10 on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Zero-Shot Object Detection with YOLO-WorldColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv9 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune RTMDet on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Segment Images with FastSAMColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLO-NAS on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Segment Images with Segment Anything Model (SAM)Colab KaggleRoboflow YouTubeGitHub arXiv
Zero-Shot Object Detection with Grounding DINOColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune DETR Transformer on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Classify Images with DINOv2Colab KaggleRoboflowGitHub arXiv
Fine-Tune YOLOv8 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv8 on Pose Estimation DatasetColab KaggleRoboflowGitHub
Fine-Tune YOLOv8 on Oriented Bounding Boxes (OBB) DatasetColab KaggleRoboflowGitHub
Fine-Tune YOLOv8 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv8 on Classification DatasetColab KaggleRoboflowGitHub
Fine-Tune YOLOv7 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv7 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune MT-YOLOv6 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv5 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv5 on Classification DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv5 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune Faster RCNN on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune SegFormer on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune ViT on Classification DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune Scaled-YOLOv4 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOS on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOR on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOX on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune ResNet34 on Classification DatasetColab KaggleRoboflow YouTube
Image Classification with OpenAI ClipColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv4-tiny Darknet on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Train a YOLOv8 Classification Model with No LabelingColab KaggleRoboflowGitHub

📍 tracker tutorials (4 notebooks)

🛠️ computer vision skills (23 notebooks)

🎬 videos

Almost every week we create tutorials showing you the hottest models in Computer Vision. 🔥 Subscribe, and stay up to date with our latest YouTube videos!

How to Choose the Best Computer Vision Model for Your Project How to Choose the Best Computer Vision Model for Your Project

Created: 26 May 2023 | Updated: 26 May 2023

In this video, we will dive into the complexity of choosing the right computer vision model for your unique project. From the importance of high-quality datasets to hardware considerations, interoperability, benchmarking, and licensing issues, this video covers it all...


Accelerate Image Annotation with SAM and Grounding DINO Accelerate Image Annotation with SAM and Grounding DINO

Created: 20 Apr 2023 | Updated: 20 Apr 2023

Discover how to speed up your image annotation process using Grounding DINO and Segment Anything Model (SAM). Learn how to convert object detection datasets into instance segmentation datasets, and see the potential of using these models to automatically annotate your datasets for real-time detectors like YOLOv8...


SAM - Segment Anything Model by Meta AI: Complete Guide SAM - Segment Anything Model by Meta AI: Complete Guide

Created: 11 Apr 2023 | Updated: 11 Apr 2023


Discover the incredible potential of Meta AI's Segment Anything Model (SAM)! We dive into SAM, an efficient and promptable model for image segmentation, which has revolutionized computer vision tasks. With over 1 billion masks on 11M licensed and privacy-respecting images, SAM's zero-shot performance is often superior to prior fully supervised results...

💻 run locally

We try to make it as easy as possible to run Roboflow Notebooks in Colab and Kaggle, but if you still want to run them locally, below you will find instructions on how to do it. Remember don't install your dependencies globally, use venv.

# clone repository and navigate to root directory
git clone git@github.com:roboflow-ai/notebooks.git
cd notebooks

# setup python environment and activate it
python3 -m venv venv
source venv/bin/activate

# install and run jupyter notebook
pip install notebook
jupyter notebook

☁️ run in sagemaker studio lab

You can now open our tutorial notebooks in Amazon SageMaker Studio Lab - a free machine learning development environment that provides the compute, storage, and security—all at no cost—for anyone to learn and experiment with ML.

Stable Diffusion Image GenerationYOLOv5 Custom Dataset TrainingYOLOv7 Custom Dataset Training
SageMakerSageMakerSageMaker

🐞 bugs & 🦸 contribution

Computer Vision moves fast! Sometimes our notebooks lag a tad behind the ever-pushing forward libraries. If you notice that any of the notebooks is not working properly, create a bug report and let us know.

If you have an idea for a new tutorial we should do, create a feature request. We are constantly looking for new ideas. If you feel up to the task and want to create a tutorial yourself, please take a peek at our contribution guide. There you can find all the information you need.

We are here for you, so don't hesitate to reach out.

Contributors

(top 30 of 41)

SkalskiP

405 commits

capjamesg

90 commits

LinasKo

48 commits

Jacobsolawetz

14 commits

roboflow/notebooks

A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-edge models like RF-DETR, YOLO11, SAM 3, and Qwen3-VL.

9,673

stars

655

commits

Jupyter Notebook

primary language

Aug 14, 2026

updated

roboflow.com/models
automatic-labeling-system
computer-vision
deep-learning
deep-neural-networks
google-colab
image-classification
image-segmentation
machine-learning
object-detection
open-vocabulary-detection
open-vocabulary-segmentation
paligemma
pytorch
qwen
tutorial
vlm
yolov5
yolov8
zero-shot-classification
zero-shot-detection

README

👋 hello

This repository offers a growing collection of computer vision tutorials. Learn to use SOTA models like YOLOv11, SAM 2, Florence-2, PaliGemma 2, and Qwen2.5-VL for tasks ranging from object detection, segmentation, and pose estimation to data extraction and OCR. Dive in and explore the exciting world of computer vision!

🚀 model tutorials (61 notebooks)

notebookopen in colab / kaggle / sagemaker studio labcomplementary materialsrepository / paper
RF-DETR Keypoint Detection and Fine-TuningColab KaggleGitHub
Object Detection with Google Gemini 3.5 FlashColab Kaggle
How to Perform OCR with GLM-OCRColab KaggleYouTubeGitHub arXiv
How to Track Objects with RF-DETR and ByteTrack TrackerColab KaggleYouTubeGitHub arXiv
Fine-Tune YOLO26 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLO26 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Segment Images with SAM3Colab KaggleRoboflow YouTubeGitHub arXiv
Segment Videos with SAM3Colab KaggleRoboflow YouTubeGitHub arXiv
Open Vocabulary Object Detection with Qwen3-VLColab KaggleGitHub
Fine-Tune RF-DETR Segmentation on Custom DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Zero-Shot Object Detection and Segmentation with Google Gemini 2.5Colab KaggleRoboflowarXiv
Fine-Tune RF-DETR on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Zero-Shot Object Detection and Segmentation with YOLOEColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv12 on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Zero-Shot Object Detection with Qwen2.5-VLColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune Qwen2.5-VL for JSON Data ExtractionColab KaggleYouTubeGitHub arXiv
Fine-Tune PaliGemma2 on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Fine-Tune PaliGemma2 for JSON Data ExtractionColab KaggleRoboflowGitHub arXiv
Fine-Tune PaliGemma2 for LaTeX OCRColab KaggleRoboflowGitHub arXiv
Fine-Tune SAM-2.1Colab KaggleRoboflow YouTubeGitHub
Fine-Tune GPT-4o on Object Detection DatasetColab KaggleRoboflow YouTube
Fine-Tune YOLO11 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLO11 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Segment Images with SAM2Colab KaggleRoboflow YouTubeGitHub arXiv
Segment Videos with SAM2Colab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune RT-DETR on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Fine-Tune Florence-2 on Object Detection DatasetColab KaggleRoboflow YouTubearXiv
Run Different Vision Tasks with Florence-2Colab KaggleRoboflow YouTubearXiv
Fine-Tune PaliGemma on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv10 on Object Detection DatasetColab KaggleRoboflowGitHub arXiv
Zero-Shot Object Detection with YOLO-WorldColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv9 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune RTMDet on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Segment Images with FastSAMColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLO-NAS on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Segment Images with Segment Anything Model (SAM)Colab KaggleRoboflow YouTubeGitHub arXiv
Zero-Shot Object Detection with Grounding DINOColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune DETR Transformer on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Classify Images with DINOv2Colab KaggleRoboflowGitHub arXiv
Fine-Tune YOLOv8 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv8 on Pose Estimation DatasetColab KaggleRoboflowGitHub
Fine-Tune YOLOv8 on Oriented Bounding Boxes (OBB) DatasetColab KaggleRoboflowGitHub
Fine-Tune YOLOv8 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv8 on Classification DatasetColab KaggleRoboflowGitHub
Fine-Tune YOLOv7 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv7 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune MT-YOLOv6 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv5 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv5 on Classification DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune YOLOv5 on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub
Fine-Tune Faster RCNN on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune SegFormer on Instance Segmentation DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune ViT on Classification DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune Scaled-YOLOv4 on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOS on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOR on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOX on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune ResNet34 on Classification DatasetColab KaggleRoboflow YouTube
Image Classification with OpenAI ClipColab KaggleRoboflow YouTubeGitHub arXiv
Fine-Tune YOLOv4-tiny Darknet on Object Detection DatasetColab KaggleRoboflow YouTubeGitHub arXiv
Train a YOLOv8 Classification Model with No LabelingColab KaggleRoboflowGitHub

📍 tracker tutorials (4 notebooks)

🛠️ computer vision skills (23 notebooks)

🎬 videos

Almost every week we create tutorials showing you the hottest models in Computer Vision. 🔥 Subscribe, and stay up to date with our latest YouTube videos!

How to Choose the Best Computer Vision Model for Your Project How to Choose the Best Computer Vision Model for Your Project

Created: 26 May 2023 | Updated: 26 May 2023

In this video, we will dive into the complexity of choosing the right computer vision model for your unique project. From the importance of high-quality datasets to hardware considerations, interoperability, benchmarking, and licensing issues, this video covers it all...


Accelerate Image Annotation with SAM and Grounding DINO Accelerate Image Annotation with SAM and Grounding DINO

Created: 20 Apr 2023 | Updated: 20 Apr 2023

Discover how to speed up your image annotation process using Grounding DINO and Segment Anything Model (SAM). Learn how to convert object detection datasets into instance segmentation datasets, and see the potential of using these models to automatically annotate your datasets for real-time detectors like YOLOv8...


SAM - Segment Anything Model by Meta AI: Complete Guide SAM - Segment Anything Model by Meta AI: Complete Guide

Created: 11 Apr 2023 | Updated: 11 Apr 2023


Discover the incredible potential of Meta AI's Segment Anything Model (SAM)! We dive into SAM, an efficient and promptable model for image segmentation, which has revolutionized computer vision tasks. With over 1 billion masks on 11M licensed and privacy-respecting images, SAM's zero-shot performance is often superior to prior fully supervised results...

💻 run locally

We try to make it as easy as possible to run Roboflow Notebooks in Colab and Kaggle, but if you still want to run them locally, below you will find instructions on how to do it. Remember don't install your dependencies globally, use venv.

# clone repository and navigate to root directory
git clone git@github.com:roboflow-ai/notebooks.git
cd notebooks

# setup python environment and activate it
python3 -m venv venv
source venv/bin/activate

# install and run jupyter notebook
pip install notebook
jupyter notebook

☁️ run in sagemaker studio lab

You can now open our tutorial notebooks in Amazon SageMaker Studio Lab - a free machine learning development environment that provides the compute, storage, and security—all at no cost—for anyone to learn and experiment with ML.

Stable Diffusion Image GenerationYOLOv5 Custom Dataset TrainingYOLOv7 Custom Dataset Training
SageMakerSageMakerSageMaker

🐞 bugs & 🦸 contribution

Computer Vision moves fast! Sometimes our notebooks lag a tad behind the ever-pushing forward libraries. If you notice that any of the notebooks is not working properly, create a bug report and let us know.

If you have an idea for a new tutorial we should do, create a feature request. We are constantly looking for new ideas. If you feel up to the task and want to create a tutorial yourself, please take a peek at our contribution guide. There you can find all the information you need.

We are here for you, so don't hesitate to reach out.

Contributors

(top 30 of 41)

SkalskiP

405 commits

capjamesg

90 commits

LinasKo

48 commits

Jacobsolawetz

14 commits

Languages

Jupyter Notebook

100.0%