14 repos
PyTorch-based vision transformer architectures and pre-trained models for image understanding and feature extraction. This cluster centers on modern computer vision models like DINOv2 and Vision Transformers (ViT), along with tools for working with model weights (safetensors format) and extracting learned image representations. Repositories here span both foundational transformer models and applications built on top of them, useful for anyone working on image classification, representation learning, or transfer learning in vision tasks.