Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-RewardNormalization-seed202 – AI Model by IDEALLab | AlphaNeural AI
You can deploy this model and start earning money today!
IDEALLab
/
Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-RewardNormalization-seed202
like
0
transformers
safetensors
qwen2
text-generation
neural-solver-synthesis
sds
grpo
ablation
reward-normalization
icml-2026
conversational
SoheylM/OpenR1-SDS-10k-seed202
Qwen/Qwen2.5-Coder-14B-Instruct
finetune
apache-2.0
text-generation-inference
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-RewardNormalization-seed202
Canonical selected checkpoint for the SDS reward-normalization ablation, seed 202.
Notes
This repo is part of the private-first staging publication pass.
It contains the canonical paper-selected checkpoint only.
Visibility is private during validation and may be changed later.
Related datasets
SoheylM/OpenR1-SDS-10k-seed202