Real-world tokens/sec numbers for 12 local coding/agentic LLMs, measured on a specific,
commonly-owned but previously unbenchmarked configuration: Apple M5 Pro, 24GB unified
memory. At the time of writing, no public benchmark existed for this exact chip + RAM
combination for these models. Includes both the official newly-open-sourced Qwen3.8-27B and
a popular uncensored/abliterated finetune of it, for comparison.… See the full description on the dataset page:
https://huggingface.co/datasets/abhisheksharma0994/apple-m5-pro-24gb-llm-tps-benchmarks.