AiEdit is a large-scale, cross-lingual speech editing dataset designed to advance research and evaluation in Speech Editing tasks. We have constructed an automated data generation pipeline comprising the following core modules:
Text Engine: Powered by Large Language Models (LLMs), this engine intelligently processes raw text to execute three types of editing operations: Addition, Deletion, and Modification.
Speech Synthesis & Editing:⦠See the full description on the dataset page: https://huggingface.co/datasets/nccm2p2/AiEdit.