Combined Fine-tuned Sesame CSM 1B Synthetic Speech Dataset
Data Format
Each sample contains:
audio: Audio array with sampling_rate (24kHz)
text: Original transcript text
speaker_id: Speaker identifier (F04, M02, FC02, MC01, F02, M04, 211, 4014)
corpus: Source corpus (TORGO, UA-Speech, LibriSpeech)
condition: Speaker condition (Dysarthric, Healthy)
model_name: Fine-tuned model name
model_type: "sesame_csm_1b_adapter"
base_model: Base model used (unsloth/csm-1b)… See the full description on the dataset page: https://huggingface.co/datasets/resproj007/sesame_pathological_synthetic_data.