We introduce PhonemeFake, a DF attack that manipulates critical speech segments using language reasoning, significantly reducing human perception and SoTA
model accuracies.
We provide example spoof audio in the viewers tab along with the transcription of the bonafide sample, the manipulated transcription and the audio timings for a small set of the data.
The dataset is split into three subsets. The spoof samples… See the full description on the dataset page:
https://huggingface.co/datasets/phonemefake/PhonemeFakeV2.