operator reproduced · 2026-10-03
Load the exact GGUF in llama.cpp and confirm a short completion.
One report, not a guarantee or failure rate.
Exact fileqwen2.5-0.5b-instruct-q4_k_m.gguf · SHA-256 74a4da8c9fdbcd15bd1f6d01d621410d31c6fc00986f5eb687824e7b93d7a9db
HardwareARM64 virtual CPU; physical model not exposed · 4.10194 GB system RAM
Runtimellama.cpp · b3991 (fc83a9e) · CPU (ARM64)
Task / outcomechat · Worked
Success criteriaThe model loaded without error and produced non-empty text before the 16-token limit.
No measurements reported.
Contributor notes
Local CPU smoke test only. It does not assess answer quality, measure speed, TTFT, memory, or establish a context limit. Prompt and generated text were not retained.