Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
phi-3-mini-4k-instruct-onnx-qnn – AI Model by llmware | AlphaNeural AI
You can deploy this model and start earning money today!
llmware
/
phi-3-mini-4k-instruct-onnx-qnn
like
0
onnx
phi3
green
llmware-chat
p3
qnn
emerald
custom_code
apache-2.0
4-bit
gptq
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
phi-3-mini-4k-onnx-qnn
phi-3-mini-4k-onnx-qnn
is an ONNX QNN int4 quantized version of
Microsoft Phi-3-mini-instruct
, providing a small fast NPU inference implementation, optimized for NPU deployment on Windows ARM64 AI PCs with Snapdragon Elite X NPU processors.
Model Description
Developed by:
microsoft
Model type:
phi3
Parameters:
3.8 billion
Model Parent:
microsoft/Phi-3-mini-instruct
Language(s) (NLP):
English
License:
Apache 2.0
Uses:
Chat, general-purpose LLM
Quantization:
int4
Backend:
qairt 2.36, ort 1.22.2, ortg 0.9
Model Card Contact
llmware on hf
llmware website