Made with Exllamav2 with default dataset.
These quants are meant to be used as a draft model for 24B Mistral models with TabbyAPI app.
8bpw version with FP16 cache probably might be the most reliable option for this purpose.
The data are Mistral's outputs and includes all kind of tasks from various datasets in English, French, German, Spanish, Italian and Portuguese. It has been trained for 2 epochs on 20k unique examples, for a total of 12 million tokens per epoch.