The Cultural and Moral Expressions in Language (CAMEL) corpus is a dataset of more than 57000 texts gathered from multiple sources including Reddit, X (Twitter), Wikipedia, and other relevant text datasets. This dataset was prepared by the Culture and Morality Lab (CaM-L) at the University of Massachusetts Amherst. For an overview of our work, visit us here 🐪
Each piece of text has been annotated by at least three human annotators. The annotators received intensive… See the full description on the dataset page:
https://huggingface.co/datasets/Culture-and-Morality-Lab/CAMEL_Dataset.