Mitsua Diffusion One is a latent text-to-image diffusion model, which is a successor of
Mitsua Diffusion CC0.
This model is
trained from scratch using only public domain/CC0 or copyright images with permission for use, with using a fixed pretrained text encoder (
OpenCLIP ViT-H/14, MIT License).
This will be used as a base model for
AI VTuber Elan Mitsua🖌️’s activity.
We are active on
a Discord server for opt-in contributors only. Communication is currently in Japanese.
This model is open access and available to all, with a Mitsua Open RAIL-M license further specifying rights and usage. The Mitsua Open RAIL-M License specifies:
All data was obtained ethically and in compliance with the site's terms and conditions.
No copyright images are used in the training of this model without the permission.
No AI generated images are in the dataset.
Approx 11M images in total with data augmentation.