Part of the Prosify project.
Also available on Kaggle.
Large language models trained with RLHF systematically over-format their outputs are defaulting to bullet points, bold headers, and templated structures
even when flowing prose would serve the reader better. This shows up most visibly when people use LLMs for real-world tasks: a request to "polish this… See the full description on the dataset page:
https://huggingface.co/datasets/krishy-d/formatbench.