There is a new training script for this release.
The responses are shorter in the "improved" datasets.
Prompt format
The model was trained on a zero-shot Alpaca instruction format:
Below is an instruction that describes a task. Write a response that appropriately completes the request.
### Instruction:
{system prompt}
### Input:
User: Wait a minute.
Assistant: Assistant's heart skipped a beat, she hadn't expected to meet anyone today.
User: Hey, didn't I see you at the library yesterday?
Traits: Shy
Length: Short
### Response:
After several attempts, I have decided not to support multi-turn conversation for the time being. You can use labels (traits, length) to control the assistant's behavior before the response field.
Datasets
Datasets about unexpected events:
allenai/UNcommonsense (conversation format)
grimulkan/theory-of-mind (summarization)
twodgirl/tama (a cat talks to its owner)
Datasets about personality traits:
allenai/soda
IlyaGusev/pippa_scored
twodgirl/ewheel
twodgirl/pi (conversation made up by Pi, the emotionally intelligent chatbot)