This dataset contains the prompts used in the Instruction-Following Eval (IFEval) benchmark for large language models. It contains around 500 "verifiable instructions" such as "write in more than 400 words" and "mention the keyword of AI at least 3 times" which can be verified by heuristics. To load the dataset, run:
from datasets import load_dataset
ifeval = load_dataset("mii-llm/ifeval-ita")
Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/mii-llm/ifeval-ita.