For test split, please use the test splits available in the following datasets:
Imadken/Lamini_formatted
Imadken/platypus_formatted
both of tests splits represents 10% of the original data
features:
- name: input
dtype: string
- name: output
dtype: string
- name: instruction
dtype: string
- name: data_source
dtype: string
- name: text
dtype: string
splits:
- name: train
num_bytes: 59540191
num_examples: 23693… See the full description on the dataset page: https://huggingface.co/datasets/Imadken/platypus_Lamini_formatted.