Large Language Models (LLMs) are transforming diverse fields and gaining increasing influence as human proxies. This development underscores the urgent need for evaluating value orientations and understanding of LLMs to ensure their responsible integration into public-facing applications. ValueBench is the first comprehensive psychometric benchmark for evaluating value orientations and value understanding in LLMs. We collect data from 44 established… See the full description on the dataset page:
https://huggingface.co/datasets/Value4AI/ValueBench.