๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Prince Canuma
@Prince_Canuma
Creator of (@Nativ_AI, mlx-audio & mlx-vlm) โ€ข working on something new Ex-@arcee_ai โ€ข @neptune_ai โ€ข
๊ฐ€์ž… July 2012
1.3K ํŒ”๋กœ์ž‰ ์ค‘    22.8K ํŒฌ
Congrats to @liquidai on LFM2.5-2.6B! Excited to have partnered with them for Day 0 support in Nativ ๐ŸŽ‰ Built for agentic + coding workflows โ€” and itโ€™s fast. On an M5 Max (48GB) with Nativ v0.2.2 โ€” full bf16, no quantization: โšก 11,231 tok/s prefill โšก ~84 tok/s decode ๐Ÿง  Full 128K context in just 8.5GB ๐Ÿ“ˆ 476 tok/s aggregate decode at batch 16 Local inference doesn't get much better than this. Get started today๐Ÿ‘‡๐Ÿฝ
๋” ๋ณด๊ธฐ