This is an anime-style text-to-image (T2I) model further trained by the Nieta.art team, based on the excellent open-source model Lumina-Image-2.0 by Alpha-VLLM. During the training process, we utilized an extensive, rich, and diverse dataset of anime images, paired with high-quality Natural Language Processing (NLP) tags, striving for the model to better understand and interpret anime-related text descriptions.
Open Testing & Spirit of Sharing:
In the spirit of technological sharing and the mutual progress of the open-source community, we have decided to release partial pre-trained weights during the extremely early stages of model training (Alpha test). We hope that this will allow interested developers and researchers to get an early touchpoint with this model for preliminary exploration and testing, and we look forward to receiving valuable feedback in the future.
⚠️ IMPORTANT NOTES AND DISCLAIMERS ⚠️
Before you download and use this model, please carefully read and understand the following points:
Extremely Early Stage: This model is currently in a very, very initial phase of development and training. This means the model is far from mature and may contain numerous known and unknown issues, defects, and unstable behaviors.
No Post-Optimization: The currently released Alpha version has NOT undergone any subsequent aesthetic alignment, style enhancement, instruction following optimization, or any other form of post-training. The raw output of the model may significantly differ from your expectations.
Extremely High Difficulty of Use: Due to the reasons above, the barrier to entry for using the current model is very high. It may require specific parameter tuning and extensive experimentation to generate specific content, and the results might not be satisfactory. Please be fully prepared for this.
Regarding Resolution Increase & Knowledge Forgetting: The model has recently undergone a training adjustment for a resolution increase and is currently in a re-adaptation phase. This may lead to more significant temporary knowledge forgetting in the short term, affecting the stability of generation results.
5. Required Components:
- Text Encoder (TE) & VAE: Directly sourced from Comfy-Org/Lumina_Image_2.0_Repackaged
- DiT Weights: This table only releases Diffusion Transformer weights (click filenames to download)
- Resources:
Do Not Share Model Weights: To ensure the controllability of the project during its development phase, please do not redistribute or share the model weight files of this Alpha version through any public or private channels.
No Authorization Granted:Crucially, the release of this Alpha test version does NOT constitute any form of usage authorization or license for the model itself, its weights, or the content generated by it. As the model is entirely unfinished in its training, its nature and capabilities are still under exploration. Nieta.art reserves all rights to the model and its derivatives. Please do not use this early test model or its outputs for any commercial purposes, official projects, or any scenarios that could lead to adverse effects.
We thank you for your interest in Nieta.art and the NietaAniLumina project. We look forward to bringing you a more complete and powerful version once the model matures.