Update 7/9/25: This model is now quantized and implemented in
this example space. Seeing preliminary VRAM usage at around ~10GB with faster inferencing. Will be experimenting with different weights and schedulers to find particularly well-performing libraries.
Highly experimental, will update with more details later.
Experimenting with FLUX.1-dev LoRAs and how it affects Kontext-dev. This model has been fused with acceleration LoRAs.
This model falls under the
FLUX.1 [dev] Non-Commercial License, please familiarize yourself with the license.