💡 Reward Modeling from Natural Language Human Feedback
This is the official dataset used in paper "Reward Modeling from Natural Language Human Feedback".
🔑 Key Features
RM-NLHF integrates multiple preference datasets.
We employ Qwen3-235B-A22B-2507 to extract key points from the human-annotated commentary portion of HelpSteer3 and reformat them into structured bullet-point lists.
In this repository, we are open-sourcing… See the full description on the dataset page: https://huggingface.co/datasets/Tongyi-ConvAI/RM-NLHF.