Llama 3.2 3B Instruct fused with the
PromptMon LoRA v2 adapter, converted to GGUF and quantized to Q4_K_M for on-device mobile inference (fllama / llama.cpp).
Use the Llama 3.2 chat template — PromptMon prompts assume it verbatim.
Inherits Llama 3.2 Community License.