ThumbLLM v0.1 Qwen3.5-4B CPU. Windows. Double-click.
On my Strix Halo (CPU only, no iGPU offload), it sits near ~20 tok/s.
Same recipe idea as the 7540U laptop post: --device none, GGUF Q4_K_M, chat + local API. Missing weights download.
Unsigned. SmartScreen will complain. Hash is on the Release. Then run anyway.
๐ GH /TeksEdge/ThumbLLM/releases/tag/thumbllm-qwen3.5-4b-mtp-q4_k_m-cpu-win-x64-v0.1.0