A collection of watermark-free Modelscope-based video models capable of generating high quality video at
448x256,
576x320 and
1024 x 576. These models were trained from the
original weights with offset noise using 9,923 clips and 29,769 tagged frames.
This collection makes it easy to switch between models with the new dropdown menu in the 1111 extension.
Simply download the contents of this repo to 'stable-diffusion-webui\models\text2video'
Or, manually download the model folders you want, along with VQGAN_autoencoder.pth.