Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
GSA-PT-Qwen2-7B-Instruct-chunk32 – AI Model by gist-sparse-attention | AlphaNeural AI
You can deploy this model and start earning money today!
gist-sparse-attention
/
GSA-PT-Qwen2-7B-Instruct-chunk32
like
0
safetensors
qwen2
gsa
gist-sparse-attention
long-context
Qwen/Qwen2-7B-Instruct
finetune
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
GSA-PT-Qwen2-7B-Instruct-chunk32
This model is
continued pretrained
from
Qwen/Qwen2-7B-Instruct
using
Gist Sparse Attention (GSA)
with chunk size
chunk32
.
Paper
GSA: Gist Sparse Attention via Learnable Compression and Selective Unfolding
Model Details
Field
Value
Base model
Qwen/Qwen2-7B-Instruct
Training type
Continued Pretraining
Chunk size
chunk32
Architecture
Qwen2-7B