This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the legacy StarCoder 7B base foundational model.
This specific partition documents the behavioral dynamics of larger-scale raw foundational weights inside automated benchmarking pipelines, establishing an anchor point to analyze… See the full description on the dataset page:
https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-mbpp-starcoder_7b.