Congrats to
@liquidai on LFM2.5-2.6B! Excited to have partnered with them for Day 0 support in Nativ 🎉
Built for agentic + coding workflows — and it’s fast.
On an M5 Max (48GB) with Nativ v0.2.2 — full bf16, no quantization:
⚡ 11,231 tok/s prefill
⚡ ~84 tok/s decode
🧠 Full 128K context in just 8.5GB
📈 476 tok/s aggregate decode at batch 16
Local inference doesn't get much better than this.
Get started today👇🏽