π₯ π₯ π₯ [08/11/2023] We release WizardMath Models.
π₯ Our WizardMath-70B-V1.0 model slightly outperforms some closed-source LLMs on the GSM8K, including ChatGPT 3.5, Claude Instant 1 and PaLM 2 540B.
π₯ Our WizardMath-70B-V1.0 model achieves 81.6 pass@1 on the GSM8k Benchmarks, which is 24.8 points higher than the SOTA open-source LLM.
π₯ Our WizardMath-70B-V1.0 model achieves 22.7 pass@1 on the MATH Benchmarks, which is 9.2 points higher than the SOTA open-source LLM.β¦ See the full description on the dataset page:
https://huggingface.co/datasets/WizardLMTeam/WizardLM_evol_instruct_V2_196k.