Local Code Arena Telemetry: MBPP Benchmark on Qwen 3.5 9B
This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the Qwen 3.5 9B foundation architecture.
This specific run establishes the mid-tier performance and throughput dynamics of a next-generation generalist instruction model on local consumer hardware.
📊 Core Performance Summary… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-mbpp-qwen3.5_9b.