Qwen3.8-27B on 1500$ oh hardware.
This is the realistic local AI experience for most people, for tons of reasons including complexity, costs, and over abundance of configs.
Still, usable. GLM-4 was about 30 tok/s at peak
The future is bright.
Goodnight friends.
顯示更多