How did I miss this?! Bonsai isnโt just an LLM family.
Back in May, PrismML released Bonsai Image 4B, a crazy low-bit version of FLUX.2 Klein 4B designed to run locally on ๐ iPhones.
Look at this
๐จ FLUX.2 Klein 4B transformer
๐พ FP16: 7.75GB
๐ณ Ternary Bonsai: 1.21GB
And the entire Apple deployment payload, including its compressed text encoder + VAE, is only:
๐ฅ 3.88GB
๐ฑ iPhone 17 Pro Max
๐ง A19 Pro / 12GB unified memory
๐ผ๏ธ 512ร512
โก 9.4 sec/image
๐ฒ MLX Swift
โ๏ธ No cloud
๐พ Bomsai Studio (App Store)
M4 Pro: ~5.8 sec/image.
No cloud.
The ~4B diffusion transformer is paired with a 4-bit Qwen3-4B text encoder, which gets unloaded after the prompt is encoded to save memory.
And thereโs also
๐ Mac / iPhone / iPad support
๐ข low-bit Gemlite builds for Nvidia GPUs
๐ Apache 2.0
This is completely separate from the Bonsai 2 27B LLM Iโve been posting about.
PrismML took a 7.75GB FLUX transformer and crunched it down to 1.21G and put image generation on an iPhone. ๐
And somehow I missed this for 4 months.