๐ Diamond Benchmark
750 real degraded speech recordings for evaluating restoration models
These are real, genuinely degraded speech recordings โ not synthetic, not
clean audio re-degraded by a pipeline. Each clip carries the damage of an actual
real-world capture: low-bitrate codec compression, narrow bandwidth, background
noise, and clipping. Together they form a fixed benchmark for measuring how
well a speech-restoration model recovers cleanโฆ See the full description on the dataset page: https://huggingface.co/datasets/nineninesix/diamond-benchmark.