This repository provides a preference dataset developed by LLM-jp, a collaborative project launched in Japan.
This dataset was created by generating the chosen response using Qwen/Qwen2.5-32B-Instruct and the rejected response using llm-jp/llm-jp-3-1.8b-instruct for the prompts in weblab-GENIAC/aya-ja-evol-instruct-calm3-dpo-masked.
This repository does not contain prompts but only the corresponding indices. Please obtain the original prompts from the original data… See the full description on the dataset page:
https://huggingface.co/datasets/llm-jp/aya-ja-evol-inst.