Benchmark results from the MoE Sovereign project — a sovereign Mixture-of-Experts AI infrastructure for regulated environments.
Dataset Contents
LLM Role Suitability Study (69 Models)
Files: results/llm_role_suitability_merged.json, results/llm_role_suitability_parallel.json
Systematic evaluation of 69 local LLMs for MoE orchestration roles (Planner, Judge, Expert).
Key findings:
61% suitable for both Planner + Judge roles
26%… See the full description on the dataset page:
https://huggingface.co/datasets/h3rb3rn/moe-sovereign-benchmarks.