This was meant to be training data to teach an LLM to do some basic document editing tasks.
Input: 150 Wikipedia articles + A request to substitute one word for another (usually a synonym)
Output: The same article, with the word substituted as requested
Format: Fastchat
Input: 224 Wikipedia articles with typos and other errors introduced randomly using the python typo library + A request to fix errors
Output:… See the full description on the dataset page:
https://huggingface.co/datasets/grimulkan/document-editing.