We provide 4 Binary-Autoencoder (BAE) tokenizers, following
Binary Latent Diffusion, with code dimension 16, 10, 24 and 32, each trained for 1,000,000 iterations with batch size 256.
The generation model architecture is adapted from
Llama2, following
LlameGen.