🤯 This entire 64GB AMD AI workstation cost its builder LESS than a single RTX 5090 and Qwen3.8-27B is hitting 111 tok/s.
This is the kind of rig that makes me want to build it because RTX 5090s are so expensive.
Setup 👇
🟠 2× Radeon AI PRO R9700 32GB
💾 64GB total dedicated VRAM
🧠 Ryzen 7500F
💾 64GB DDR5
🔌 PCIe 5.0 x8 per GPU
⚡ vLLM Radiance + TP2 + MTP
Running Qwen3.8-27B Quark AWQ MXFP4 ...
🚀 111.4 tok/s median decode
⚡ 4,224–4,410 tok/s prefill
⏱️ 81ms TTFT
📚 131K server context
Native FP8 still managed 87.6 tps
🎮 RTX 5090 → 32GB VRAM
🟠 2× R9700 → 64GB VRAM
The builder says the entire machine cost ~€4,000, more than €1,000 less than a single RTX 5090 in his market.
So for less money, he ended up with 2× the VRAM + 111 tok/s on Qwen3.8-27B. 🔥
Reddit Link in ALT