Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
Qwen3.5-4B-MiniFantasy-GGUF – AI Model by nuofang | AlphaNeural AI
You can deploy this model and start earning money today!
nuofang
/
Qwen3.5-4B-MiniFantasy-GGUF
like
0
gguf
llama.cpp
quantized
imatrix
Nubinu/Qwen3.5-4B-MiniFantasy
quantized
endpoints_compatible
us
conversational
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Auto-Quantized GGUF Model
This repository contains automated GGUF quantization files for
Nubinu/Qwen3.5-4B-MiniFantasy
.
The calibration data for the imatrix is targeted at Chinese novels and role-playing (RP), while preserving logic and common sense.
imatrix 的校准数据以中文的小说、角色扮演为目标,同时保留逻辑和常识。
📊 Perplexity Evaluation
(Tested against the provided calibration dataset)
Base (F16/BF16)
: PPL = 16.5300 +/- 0.13828
IQ4_XS
: PPL = 14.1688 +/- 0.11587
Q4_K_M
: PPL = 14.0570 +/- 0.11458
Q5_K_M
: PPL = 13.9765 +/- 0.11426