Quantization made by Richard Erkhov.
Apache 2.0, following the TinyLlama base model.
Hardware: 1*RTX A5000, ~16 hours to complete 1 epoch. GPU from autodl.com, cost around $3 for this finetuning.
https://wandb.ai/jeff200402/TinyLlama-Orca?workspace= for more details.