Views
No views yet
audio to audio simulation using so-vits-4.1, as training the model is resource intensive, but not so much for infering an audio. Every model included has been trained for at least 20K stepsso-vits-svc-4.1
│
├───configs
│ ├───config.json - config file for default training
│ └───diffusion.yaml - config file for diffusion training
│
└───logs
└───44k
├───G_(name of character).pth - Default model
├───(name of character)Kmeans.pt - fusion model
└───diffusion
└───(name of character).pt - difussion model for character
data_set - dataset used for training, audio cut to slices.LICENSE.txt for more information.
If used, please attatch link to the repo.