Register and share your invite link to earn from video plays and referrals.

KC
@karanC_12
AI â€ĸ Models â€ĸ AGI â€ĸ Open Source
Joined November 2022
335 Following    1.1K Followers
This should not be possible. Qwen3.8-27B just hit 70 tokens per second on a MacBook Pro M5 Max. Same quality. Up to 4.6× faster than normal decoding. A frontier-level open model running this fast on a laptop. Local AI just became actually usable. 👀
Show more
DFlash 2 is here! Qwen3.8-27B at 70 tok/s on an M5 Max MacBook Pro. ⚡ Up to 4.6× the speed of autoregressive decoding, with the same output. This is the next generation of DFlash, seeded at Z Lab and upgraded at Inco AI. Get one more accepted token on every pass, for free!
Show more