19 repos
Semantic segmentation models built on the SegFormer architecture, a transformer-based approach for pixel-level image classification. The cluster centers on fine-tuned SegFormer variants (B0, B2, B5) trained on standard datasets like ADE20K and Cityscapes, demonstrating different model scales and resolution configurations. Repositories here provide pre-trained checkpoints, inference implementations, and documentation for applying efficient vision transformers to dense prediction tasks in computer vision.