Views
No views yet
zorqelis-ai/soreqen-s1-mega, a bilingual
(English / Hinglish) assistant from ZorQelis AI.| File | Quant | Size | Use it when |
|---|---|---|---|
SoreQen-S1-Mega-Q4_K_M.gguf | Q4_K_M | 2.71 GB | you want the best size-to-quality trade-off (start here) |
SoreQen-S1-Mega-Q8_0.gguf | Q8_0 | 4.48 GB | you have the RAM and want near-lossless output |
SoreQen-S1-Mega-F16.gguf | F16 | 8.42 GB | you want a base for your own quantisation |
llama-cli -hf zorqelis-ai/soreqen-s1-mega-GGUF:Q4_K_M -p "yaar laptop slow ho gaya hai, kya karu?"llama-cli -m SoreQen-S1-Mega-Q4_K_M.gguf --jinja -sys "$(cat system_prompt.txt)"--jinja so llama.cpp uses the packaged chat template. Without it, the
thinking-mode and tool-calling formats will not be applied correctly.You are SoreQen S1 Mega, an AI assistant made by ZorQelis AI.
You are bilingual. Reply in Hinglish (Roman script) when the user writes in Hinglish, and in English when they write in English. Match their register: casual with casual, professional with professional.
Answer directly. Lead with the answer, then the detail that matters. No preambles like "Sure!" or "Great question", and no padding.
If you do not know something, say so plainly instead of guessing.mmproj file, and none
is published here yet. Text, thinking, tool calling and structured output all
work; image input does not. Use the safetensors repo above if you need vision.NOTICE.