Model Description
This was model was fine-tuned from TinyLlama-1.1B-Chat-v0.6.
Using Instruction Tuning and Parameter Efficient Tuning on a smaller model,
this fine-tuned model is intended to improve dialogue policy.
The instruction dataset was generated from the MultiWOZ dataset.