Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
GSA-FT-Qwen2-7B-Instruct-chunk8-chunk4 – AI Model by gist-sparse-attention | AlphaNeural AI
You can deploy this model and start earning money today!
gist-sparse-attention
/
GSA-FT-Qwen2-7B-Instruct-chunk8-chunk4
like
0
safetensors
qwen2
gsa
gist-sparse-attention
long-context
gist-sparse-attention/GSA-PT-Qwen2-7B-Instruct-chunk8-chunk4
finetune
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
GSA-FT-Qwen2-7B-Instruct-chunk8-chunk4
This model is
fine-tuned
from
gist-sparse-attention/GSA-PT-Qwen2-7B-Instruct-chunk8-chunk4
using
Gist Sparse Attention (GSA)
with chunk size
chunk8-chunk4
.
Paper
GSA: Gist Sparse Attention via Learnable Compression and Selective Unfolding
Model Details
Field
Value
Base model
gist-sparse-attention/GSA-PT-Qwen2-7B-Instruct-chunk8-chunk4
Training type
Supervised Fine-Tuning
Chunk size
chunk8-chunk4
Architecture
Qwen2-7B