The model was evaluated using a Google Colab instance with a free T4 GPU. To accommodate the VRAM constraints of the hardware while maintaining inference fidelity, I utilized the transformers library alongside bitsandbytes to load the model using 8-bit quantization.
Here is the exact code… See the full description on the dataset page: https://huggingface.co/datasets/Dimeji12/Dimeji-Fatima-Fellowship.