๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

David Dalcu
@ddalcu
I build things, local LLM engines & UIโ€™s. Created for Desktop and MLX Chat for iOS. (Zig based engine)
๊ฐ€์ž… June 2010
583 ํŒ”๋กœ์ž‰ ์ค‘    878 ํŒฌ
Does anyone want to attempt to squeeze more performance from a M5 Max & Qwen Flash Next ? Here is the plan: Please submit a PR, and I will code review it and merge it. I don't have a M5Max, otherwise I would do it obviously... I heard @ivanfioravanti likes challenges like this ๐Ÿ˜…
๋” ๋ณด๊ธฐ
@ddalcu Did exactly the same test as you but with Macbook Pro M5 max. Really cool, now I will connect it to Hermes and try some more. Thanks for fantastic work!