Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Llama-3.1-Swallow-8B-Instruct-v0.5-gguf – AI Model by marcelone | AlphaNeural AI
You can deploy this model and start earning money today!
marcelone
/
Llama-3.1-Swallow-8B-Instruct-v0.5-gguf
like
0
quantized
tokyotech-llm/Llama-3.1-Swallow-8B-Instruct-v0.5
conversational
en
endpoints_compatible
gguf
template
ja
llama3.1
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Swallow-8B-it-v05-gguf-q6_k-mixed-v1
Quantization Type
: Mixed Precision (
q5_K
,
q6_K
,
q8_0
)
Bits Per Weight (BPW)
:
7.13
Swallow-8B-it-v05-gguf-q6_k-mixed-v2
Quantization Type
: Mixed Precision (
q6_K
,
q8_0
)
Bits Per Weight (BPW)
:
7.50
Swallow-8B-it-v05-gguf-q8_0-mixed-v1
Quantization Type
: Mixed Precision (
bf16
,
q4_K
,
q5_K
,
q6_K
,
q8_0
)
Bits Per Weight (BPW)
:
8.01
Swallow-8B-it-v05-gguf-q8_0-mixed-v2
Quantization Type
: Mixed Precision (
bf16
,
q5_K
,
q6_K
,
q8_0
)
Bits Per Weight (BPW)
:
9.31
Swallow-8B-it-v05-gguf-q8_0-mixed-v3
Quantization Type
: Mixed Precision (
bf16
,
q6_K
,
q8_0
)
Bits Per Weight (BPW)
:
11.44
Swallow-8B-it-v05-gguf-q8_0-mixed-v4
Quantization Type
: Mixed Precision (
bf16
,
q8_0
)
Bits Per Weight (BPW)
:
13.38