Self-Supervised Vision Model Training

12 repos

Training and analysis of self-supervised learning models for computer vision, particularly focusing on DINO (Emerging Properties in Self-Supervised Vision Transformers) and related approaches. These repositories contain implementations, experiments, and ablations of vision transformer models trained without labeled data, exploring different architectural configurations, patch aggregation strategies, and residual connections to understand emergent properties in self-supervised visual representations.