Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
openthoughts_0.3k_eval_2e29 – Dataset by mlfoundations-dev | AlphaNeural AI
You can deploy this model and start earning money today!
mlfoundations-dev
/
openthoughts_0.3k_eval_2e29
like
0
1K<n<10K
parquet
tabular
text
datasets
dask
mlcroissant
polars
us
Views
No views yet
Model card
Files and Versions
Community
API
mlfoundations-dev/openthoughts_0.3k_eval_2e29
Precomputed model outputs for evaluation.
Evaluation Results Summary
Metric AIME24 AMC23 MATH500 MMLUPro JEEBench GPQADiamond LiveCodeBench CodeElo CodeForces AIME25 HLE LiveCodeBenchv5
Accuracy 15.7 53.2 71.4 30.0 39.5 38.2 27.7 5.1 6.2 12.0 1.6 17.3
AIME24
Average Accuracy: 15.67% ± 2.79% Number of Runs: 10
Run Accuracy Questions Solved Total Questions
1 0.00% 0 30
2… See the full description on the dataset page:
https://huggingface.co/datasets/mlfoundations-dev/openthoughts_0.3k_eval_2e29
.