Swin Transformer model pre-trained on Color Multi Fractal DB 1k (1 million images, 1k classes) at resolution 224x224 for 300 epochs, developed by
ELAN MITSUA Project / Abstract Engine.
This model is trained exclusively on 1 million fractal images which relies solely on mathematical formulas, so no real images or pretrained models are used for this training.
The Swin Transformer is a type of Vision Transformer and can be utilized for various downstream tasks.
It was introduced in the paper
Swin Transformer: Hierarchical Vision Transformer using Shifted Windows by Liu et al. and first released in
this repository.