Intel's 32GB Arc Pro B70 finally got a proper Qwen3.8-27B Local AI workout.
So this is why the 32GB VRAM is important, even from an Intel GPU.
Here is Gigazine's setup. Windows 11 using Unsloth Desktop + llama.cpp ...
๐ง Qwen3.8-27B Q4_K_XL
๐พ 30.0GB VRAM + ~1GB shared RAM
๐ 23.4 tps
Then they tried Q8:
๐ง Qwen3.8-27B Q8_K_XL
๐พ 30.2GB VRAM + ~1.5GB shared RAM
โก 15.9 tps
That's a full 27B-class model running at very usable speeds on an Intel GPU.
And also tested text-to-image models.
๐จ Z-Image-Turbo
1024ร1024 / 8 steps
โก๏ธ ~5.3 sec/image after warmup
๐ฅ MiniMax H3 video also ran
โฆbut these are much slower, demonstrating where Nvidia's more mature AI software stack still matters.
The hardware setup...
๐ฎ Arc Pro B70
๐พ 32GB GDDR6
โก 608 GB/s bandwidth
๐ 230W
And here's the interesting value angle.
In Japan, the tested ASRock B70 was selling for:
๐ฐ ยฅ298,054
while many RTX 5090s were selling above:
๐ฐ ยฅ900,000
โ ๏ธ That's Japan-specific street pricing, NOT a universal B70-vs-5090 price comparison.
But THIS is why the B70 interests me for Local AI.
It won't beat a 5090 but it's giving you 32GB of VRAM at a much lower entry point.
23.4 tok/s on Qwen3.8-27B Q4 good enough for you?
๐ GIGAZINE review in ALT