This model is a SwinV2-Large (384x384) vision transformer trained for image classification on Jet-colored gradient maps. The model learns to identify visual patterns in synthetic or colormap-encoded data to be suitable for detecting GAN generated images.