This repository contains a Harzva-produced GGUF conversion of Qwen/Qwen3-0.6B for mobile llama.cpp runtimes.
This model is maintained as part of the Mobile Model Playground, which records conversion provenance, checksums, compatibility, and mobile validation evidence.
Status: conversion candidate · Android emulator contract checked. No physical-phone performance claim is made yet.
Distribution
Hugging Face — international discovery and direct HTTPS download.
Environment class: macOS arm64 host, Metal backend
Prompt: Q: What is 2+2? A:
Response: 2 + 2 = 4.
Prompt processing: 952.8 tok/s
Generation: 112.2 tok/s
This only verifies that the converted artifact loads and generates a sane answer on the host.
Android ARM64 emulator
Model load: 1,389 ms
First token: 8,026 ms
Generation: 0.289 tok/s
Artifact file size reported by the test: 456 MB
These numbers came from a low-performance Android ARM64 AVD. They are contract evidence, not phone performance. Peak PSS, temperature, power, sustained stability, and physical-device quality remain unverified.
Sanitized machine-readable records for the host, Android AVD, and GitCode post-publication checks are available under evidence/runs/. Hugging Face publication verification is tracked independently by the Mobile Model Playground so one distribution endpoint cannot inherit another endpoint's evidence.