Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
t1-enhanced-eval-hosted_vllm-qwen-qwen3-1-7b – Dataset by reasoning-degeneration-dev | AlphaNeural AI
You can deploy this model and start earning money today!
reasoning-degeneration-dev
/
t1-enhanced-eval-hosted_vllm-qwen-qwen3-1-7b
like
0
mit
n<1K
parquet
optimized-parquet
text
datasets
pandas
polars
mlcroissant
us
t1_synthesize_knowledge
enhanced_evaluation
countdown
comparison
Views
No views yet
Model card
Files and Versions
Community
API
t1-enhanced-eval-hosted_vllm-qwen-qwen3-1-7b
Enhanced evaluation results comparing base model performance against performance with synthesized domain knowledge injected into the prompt.
Performance Comparison
Metric Base Enhanced Delta
pass@1 0.2800 (28.0%) 0.4000 (40.0%) +0.1200 (+12.0%)
pass@2 0.3600 (36.0%) 0.4900 (49.0%) +0.1300 (+13.0%)
pass@3 0.4000 (40.0%) 0.5300 (53.0%) +0.1300 (+13.0%)
pass@4 0.4700 (47.0%) 0.5700 (57.0%) +0.1000 (+10.0%)… See the full description on the dataset page:
https://huggingface.co/datasets/reasoning-degeneration-dev/t1-enhanced-eval-hosted_vllm-qwen-qwen3-1-7b
.