A benchmark for master data governance reasoning.
Enterprises are pointing language models at their own operational data and asking
them to make governance decisions: is this a duplicate supplier, does this record
breach the data contract, what breaks if we retire this reference value, can this
change be auto-approved. These decisions are expensive when they are wrong, and
they are not covered by any existing public benchmark.
MDGov-Bench is 250 items across five… See the full description on the dataset page:
https://huggingface.co/datasets/Babita11/mdgov-bench.