Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
gsm8k-qwen3-0.6B-rollouts – Dataset by RLHF-Book | AlphaNeural AI
You can deploy this model and start earning money today!
RLHF-Book
/
gsm8k-qwen3-0.6B-rollouts
like
0
text-generation
en
mit
1K<n<10K
parquet
text
datasets
dask
polars
mlcroissant
us
Views
No views yet
Model card
Files and Versions
Community
API
GSM8K Qwen3-0.6B Rollouts
Verifier-labeled solutions sampled from Qwen/Qwen3-0.6B with vLLM 0.26.0 and its pytorch top-k/top-p sampler for every prompt in the train and test splits of openai/gsm8k.
Dataset size and label balance
train: 7,473 prompt rows; 747,300 rollouts; 602,611 correct (80.64%), 144,689 incorrect (19.36%) test: 1,319 prompt rows; 131,900 rollouts; 102,245 correct (77.52%), 29,655 incorrect (22.48%)
Each row contains one source prompt and 100… See the full description on the dataset page:
https://huggingface.co/datasets/RLHF-Book/gsm8k-qwen3-0.6B-rollouts
.