๐ฐ A 1.56GB local model just took the #
1# spot for open models under 4B on Artificial Analysis. Let's run it on ThumbLLM.
OpenBMB released MiniCPM5-2B TODAY.
Stats ๐
๐ง 2.52B dense parameters
๐พ Q4_K_M GGUF: just 1.56GB
๐ 131K native context
โ๏ธ Apache 2.0
๐ฆ llama.cpp
๐ข Ollama
๐ฅ๏ธ LM Studio
๐ MLX
๐ค Native tool calling + agent training
Artificial Analysis Intelligence Index v4.2 ...
๐ฅ MiniCPM5-2B โ 15
Qwen3.5-4B Reasoning โ 14*
Qwen3.5-9B Reasoning โ 15*
Granite 4.2 3B โ 11
*Qwen scores are estimated by Artificial Analysis.
So a 1.56GB Q4 GGUF is landing in the same AA Intelligence Index tier as Qwen3.5-9B.
๐ค GDPval-AA v2 Elo โ 831
๐ฆ ฯยณ-Banking โ 21%
โก 19K output tokens/task vs 56K for Ling 3.0 Tiny
OpenBMB even released the training data and an official DSpark speculative decoding model.
No trustworthy local tok/s numbers yet.
That is the benchmark I want next. ๐ฅ Let's run it on ThumbLLM!
๐ HF /openbmb/MiniCPM5-2B