Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
t1-distill-aime-397b-to-30b-instruct-enhanced – Dataset by reasoning-degeneration-dev | AlphaNeural AI
You can deploy this model and start earning money today!
reasoning-degeneration-dev
/
t1-distill-aime-397b-to-30b-instruct-enhanced
like
0
mit
n<1K
parquet
optimized-parquet
text
datasets
pandas
polars
mlcroissant
us
strategy_distillation
aime
enhanced
reasoning_malleability
sticky_reasoning
Views
No views yet
Model card
Files and Versions
Community
API
t1-distill-aime-397b-to-30b-instruct-enhanced
Part of the Strategy Distillation experiment exploring whether synthesized reasoning strategies from large models can boost smaller models on AIME math.
Results
pass@1: 30% (3/10) Original 10 AIME questions, with distilled strategy from 397B traces.
Full Experiment Results
Condition pass@1
Instruct base (original, n=10) 50% (5/10)
Instruct + strategy (original, n=10) 30% (3/10)
Thinking base (original… See the full description on the dataset page:
https://huggingface.co/datasets/reasoning-degeneration-dev/t1-distill-aime-397b-to-30b-instruct-enhanced
.