Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
qwen3-4b-hard-2048-lr1e-3-merged-fp16 – AI Model by yuk1chan | AlphaNeural AI
You can deploy this model and start earning money today!
yuk1chan
/
qwen3-4b-hard-2048-lr1e-3-merged-fp16
like
0
transformers
safetensors
qwen3
text-generation
sft
qlora
unsloth
structured-output
conversational
en
daichira/structured-hard-sft-4k
Qwen/Qwen3-4B-Instruct-2507
finetune
apache-2.0
text-generation-inference
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
qwen3-4b-hard-2048-lr1e-3-merged-fp16
This model is a merged FP16 version of Qwen/Qwen3-4B-Instruct-2507 fine-tuned with SFT.
Training Configuration
Dataset: daichira/structured-hard-sft-4k
Max sequence length: 2048
Learning rate: 1e-3
Epochs: 1
Method: QLoRA (SFT), merged to FP16