Register and share your invite link to earn from video plays and referrals.

David Hendrickson
@TeksEdge
CEO & Founder | PhD | Startup Advisor | @Columbia | Author Generative Software Engineering | ๐Ÿ”” Follow for AI & Vibe Coding Tips ๐Ÿ‘‡
Joined July 2023
549 Following    11.2K Followers
How did I miss this?! Bonsai isnโ€™t just an LLM family. Back in May, PrismML released Bonsai Image 4B, a crazy low-bit version of FLUX.2 Klein 4B designed to run locally on ๐ŸŽ iPhones. Look at this ๐ŸŽจ FLUX.2 Klein 4B transformer ๐Ÿ’พ FP16: 7.75GB ๐ŸŒณ Ternary Bonsai: 1.21GB And the entire Apple deployment payload, including its compressed text encoder + VAE, is only: ๐Ÿ”ฅ 3.88GB ๐Ÿ“ฑ iPhone 17 Pro Max ๐Ÿง  A19 Pro / 12GB unified memory ๐Ÿ–ผ๏ธ 512ร—512 โšก 9.4 sec/image ๐Ÿ“ฒ MLX Swift โ˜๏ธ No cloud ๐Ÿ’พ Bomsai Studio (App Store) M4 Pro: ~5.8 sec/image. No cloud. The ~4B diffusion transformer is paired with a 4-bit Qwen3-4B text encoder, which gets unloaded after the prompt is encoded to save memory. And thereโ€™s also ๐ŸŽ Mac / iPhone / iPad support ๐ŸŸข low-bit Gemlite builds for Nvidia GPUs ๐Ÿ”“ Apache 2.0 This is completely separate from the Bonsai 2 27B LLM Iโ€™ve been posting about. PrismML took a 7.75GB FLUX transformer and crunched it down to 1.21G and put image generation on an iPhone. ๐Ÿ‘€ And somehow I missed this for 4 months.
Show more