Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
SFT-EN-01-29-2026 – AI Model by drewli20200316 | AlphaNeural AI
You can deploy this model and start earning money today!
drewli20200316
/
SFT-EN-01-29-2026
like
0
tensorboard
safetensors
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
SFT English Medical Model - Qwen3-4B
Overview
Base Model: Qwen3-4B
Training: DeepSpeed-Chat SFT with LoRA
Dataset: UltraMedical English (9K train, 1K eval)
Date: 2026-01-29
Training Config
LoRA dim: 64
Learning rate: 2e-5
Batch size: 2
Gradient accumulation: 4
ZeRO stage: 2
Dtype: bf16
Results
Final PPL: 2.498
Final Loss: 0.915
Directory
model/ - SFT model weights
data/ - Training data
scripts/ - Training scripts
code/ - Modified DeepSpeed-Chat code