This dataset contains deepfakes in Brazilian Portuguese created with XTTS model.
The dataset was created using the XTTS model, which is a Text-to-Speech model pre-trained in several languages including Portuguese.
In order to generate the mentioned deepfakes, the model was fed with recordings from the CETUC Corpus,
made available by Fala Brasil Group. It contains speeches from 101 speakers, totaling 140 hours of audio.… See the full description on the dataset page:
https://huggingface.co/datasets/unfake/fake_voices.