๐ฅ How fast does my MacBook Pro M5 Max run Qwen3.8-27B?
๐ There's a new benchmarking tool and leaderboard for that.
๐ฏ Liquid AI + Artificial Analysis released Pipette, an open-source benchmark built for AI running on local hardware.
๐ Pipette benchmarks the inference stack, including ...
๐ง Model
๐๏ธ Quantization
โ๏ธ Runtime
๐ป Device
๐ Context length
So instead of wondering how fast Qwen3.8 is, now you can find out how fast Qwen Q4 is in llama.cpp on my MacBook Pro M5 Max (or other local AI machines).
According to AA and Liquid, the dataset currently has ...
๐ 10,000+ benchmark results
๐งช 1,000+ tested configurations
โ model ร quantization ร runtime ร device ร context
๐ค ~35 model classes
๐๏ธ 7 quantization levels
โ๏ธ llama.cpp
๐ฑ iPhone + Android
๐ป Mac + Windows PCs
๐ง Strix Halo results coming soon
And it measures what matters to Local AI users...
โก Decode tps
๐ Prefill speed
โฑ๏ธ Latency
๐พ Peak memory
๐ง Model quality
โ ๏ธ Caveat, current benchmark only covers ...
โ MacBook Pro โ M5 Max
โ iPhone 17 Pro
โ Samsung Galaxy S26 Ultra
๐ pipette dot liquid dot ai