Fun-CineForge contains an end-to-end dataset pipeline for producing large-scale dubbing datasets and an MLLM-based dubbing model designed for diverse cinematic scenes.
Using this pipeline, we constructed the first large-scale Chinese television dubbing dataset CineDub-CN, which includes rich annotations and diverse scenes.
In monologue, narration, dialogue, and… See the full description on the dataset page:
https://huggingface.co/datasets/FunAudioLLM/CineDub-Example.