The vietdata/mixed-llm-instruction dataset is an open-source collection designed for instruction tuning and prompt recovery. This dataset comprises three key columns: prompt, context, and response. Prompts and contexts are sourced from the databricks/databricks-dolly-15k dataset. We further use LLMs to generate rewriting prompts (change stype, tone, etc.). Each rewrite prompt is paired with a randomly selected context from the… See the full description on the dataset page: https://huggingface.co/datasets/vietdata/mixed-llm-instruction.