Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
TinyLlama-1.1B-compressed-tensors-kv-cache-scheme – AI Model by nm-testing | AlphaNeural AI
You can deploy this model and start earning money today!
nm-testing
/
TinyLlama-1.1B-compressed-tensors-kv-cache-scheme
like
0
transformers
safetensors
llama
text-generation
autotrain_compatible
text-generation-inference
endpoints_compatible
compressed-tensors
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
This model is outdated and uses very old kv-cache quant scheme. It should be removed from HF-Hub once we are certain it is not used anywhere else. I have removed it from vLLM's CI in this PR:
https://github.com/vllm-project/vllm/pull/30141