Why Qwen3 0.6B is in The Hub
The smallest matched X1S model completed both CPU runtimes and all three explicit llama.cpp device modes.
It was small enough to show the CPU, mixed host-operation, and full-Vulkan differences without crossing into the larger-model i915 failures.
The rows on this page use Q4_K_M through llama.cpp 9a286ac and Ollama 0.32.1. They describe those exact artifacts and runs, not every version that shares the model name.