Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
phi-4-4bit-bnb – AI Model by fhamborg | AlphaNeural AI
You can deploy this model and start earning money today!
fhamborg
/
phi-4-4bit-bnb
like
0
transformers
safetensors
phi3
text-generation
phi
phi4
nlp
math
code
chat
conversational
custom_code
en
microsoft/phi-4
quantized
mit
autotrain_compatible
text-generation-inference
endpoints_compatible
4-bit
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Phi-4 GPTQ (4-bit Quantized)
Model
Model Description
This is a
4-bit quantized
version of the Phi-4 transformer model, optimized for
efficient inference
while maintaining performance.
Base Model
:
Phi-4
Quantization
: bnb (4-bit)
Format
:
safetensors
Tokenizer
: Uses standard
vocab.json
and
merges.txt
Intended Use
Fast inference with minimal VRAM usage
Deployment in resource-constrained environments
Optimized for
low-latency text generation
Model Details
Attribute
Value
Model Name
Phi-4 GPTQ
Quantization
4-bit (GPTQ)
File Format
.safetensors
Tokenizer
phi-4-tokenizer.json
VRAM Usage
~X GB (depending on batch size)