Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
Qwen3-4B-MedCaseReasoning-RL – AI Model by nsk7153 | AlphaNeural AI
You can deploy this model and start earning money today!
nsk7153
/
Qwen3-4B-MedCaseReasoning-RL
like
0
safetensors
qwen3
medical
reinforcement-learning
healthcare
Qwen/Qwen3-4B-Instruct-2507
finetune
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Qwen3-4B-MedCaseReasoning-RL
Qwen3-4B fine-tuned with RL on MedCaseReasoning for clinical case analysis. LoRA weights properly merged.
Model Details
Base Model
:
Qwen/Qwen3-4B-Instruct-2507
Training Method
: Reinforcement Learning (GRPO) with LoRA
Framework
:
verifiers
+
prime-rl
Usage
Please ask your administrator.
License
Apache 2.0