This dataset is a small, synthetic but theory-grounded benchmark designed to support human-centered evaluation of AI systems under uncertainty.
It accompanies the human_ai_trust metric in Hugging Face Evaluate.
Each row represents a single human–AI interaction instance with the following fields:
prediction: model prediction (binary)
reference: ground truth label
confidence: model confidence in… See the full description on the dataset page:
https://huggingface.co/datasets/Dyra1204/human_ai_trust_demo.