Repo:roleplaiapp/AceInstruct-7B-Q4_K_S-GGUF Original Model:AceInstruct-7BOrganization:nvidiaQuantized File:aceinstruct-7b-q4_k_s.ggufQuantization:GGUFQuantization Method:Q4_K_S Use Imatrix:False Split Model:False
Overview
This is an GGUF Q4_K_S quantized version of AceInstruct-7B.
Quantization By
I often have idle A100 GPUs while building/testing and training the RP app, so I put them to use quantizing models.
I hope the community finds these quantizations useful.