This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the next-generation StarCoder2 3B base foundational model.
This specific partition documents the behavioral dynamics of modern raw foundational weights inside automated conversational pipelines, highlighting the persistent… See the full description on the dataset page:
https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-starcoder2_3b.