Vision Transformer Models & Feature Extraction

14 repos

PyTorch-based vision transformer architectures and pre-trained models for image understanding and feature extraction. This cluster centers on modern computer vision models like DINOv2 and Vision Transformers (ViT), along with tools for working with model weights (safetensors format) and extracting learned image representations. Repositories here span both foundational transformer models and applications built on top of them, useful for anyone working on image classification, representation learning, or transfer learning in vision tasks.

Python · 1
pytorch ·797
image-feature-extraction ·793
safetensors ·793
vision ·789
transformers ·771
pathology ·711
histology ·700
foundation-model ·637
foundation-models-for-pathology ·633
deep-learning ·633