Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
a1_math_metamath_eval_1331 – Dataset by mlfoundations-dev | AlphaNeural AI
You can deploy this model and start earning money today!
mlfoundations-dev
/
a1_math_metamath_eval_1331
like
0
1K<n<10K
parquet
tabular
text
datasets
dask
mlcroissant
polars
us
Views
No views yet
Model card
Files and Versions
Community
API
mlfoundations-dev/a1_math_metamath_eval_1331
Precomputed model outputs for evaluation.
Evaluation Results Summary
Metric AIME24 AMC23 MATH500 GPQADiamond MMLUPro LiveCodeBench CodeElo JEEBench
Accuracy 13.7 58.0 74.2 39.6 28.8 9.8 2.5 34.4
AIME24
Average Accuracy: 13.67% ± 1.29% Number of Runs: 10
Run Accuracy Questions Solved Total Questions
1 13.33% 4 30
2 16.67% 5 30
3 13.33% 4 30
4 13.33% 4 30
5 13.33% 4 30… See the full description on the dataset page:
https://huggingface.co/datasets/mlfoundations-dev/a1_math_metamath_eval_1331
.