Views
No views yet
| Property | Value |
|---|---|
| Parameters | 1,147,766,784 (~1.15B) |
| Layers | 24 |
| Attention heads | 16 (GQA: 4 KV heads) |
| Hidden dim | 2048 |
| FFN | SwiGLU |
| Normalization | RMSNorm |
| Position | RoPE with YaRN context extension |
| Context | 4,096 tokens (extensible via YaRN) |
| Tokenizer | Custom SentencePiece, 32K vocab |
| Architecture | Mini-Llama style: RoPE, GQA, SwiGLU, RMSNorm, Flash Attention |
| Model | Params | Status |
|---|---|---|
| Orion Atlas 1B | 1.15B | 🟡 Training |
| Orion Atlas 3B | ~3B | 📋 Planned |
| Orion Atlas 7B | 8.77B | 🏗️ Architecture released |
| Orion Atlas 14B | ~14B | 📋 Planned |
| Orion Atlas 37B | ~37B | 📋 Planned |
model.py for inference code.# Coming with weights release1@misc{palermini2026orionatlas,
2 title={Orion Atlas: A Mamba-2 Hybrid Architecture with Differential Attention for Agentic Language Models},
3 author={Avery Palermini},
4 year={2026},
5 institution={Cendrix AI},
6}