Register and share your invite link to earn from video plays and referrals.

David Hendrickson
@TeksEdge
CEO & Founder | PhD | Startup Advisor | @Columbia | Author Generative Software Engineering | ๐Ÿ”” Follow for AI & Vibe Coding Tips ๐Ÿ‘‡
544 Following    8.8K Followers
๐Ÿš€ Holy ๐Ÿ’ฉ! Major Local AI Breakthrough! ๐Ÿง  744B-parameter GLM-5.2 (1.5 TB total) is now running on just ~25 GB RAM โ€” no discrete GPU required! ๐Ÿ‘€ Wut!? Must test!! Italian engineer @JustVugg built Colibrรฌ, a pure C inference engine (single ~2.4k line file, zero runtime deps) that: โ€ข ๐Ÿ›ก๏ธ Keeps the dense core (~10 GB at int4) resident in RAM โ€ข ๐Ÿ“€ Streams 21,504+ MoE experts from fast NVMe on demand (only ~40B active per token) โ€ข โšก Supports native MTP speculative decoding + MLA attention Result: Frontier-class model on everyday consumer hardware! ๐Ÿ“Š Current speeds: โ€ข 25 GB RAM setup โ†’ 0.05โ€“0.1 tok/s (disk-bound) โ€ข Higher RAM + fast SSD โ†’ up to 1+ tok/s (warm) โ€ข It's a start... what could you do with a 5090? ๐Ÿ’ก Big opportunity: Pair it with Phison aiDAPTIV+ AI SSDs to kill the I/O bottleneck ๐Ÿ‘‰ smarter caching, prefetching & KV offload could make it dramatically faster! This is a huge step toward truly accessible local frontier AI. ๐Ÿ”— GitHub: JustVugg/colibri
Show more
๐Ÿš€ Local Inferencing Upgrade: AMDโ€™s next high-end APU is in development Codenamed Medusa Halo (expected as Ryzen AI Max 500 series): โ€ข Zen 6 CPU cores โ€ข RDNA 5 graphics โ€ข LPDDR6 memory support Leaks suggest it could deliver ~80% higher memory bandwidth than the current Strix Halo (Ryzen AI MAX+ 395), thanks to faster LPDDR6. This should bring meaningful improvements in: ๐ŸŽฎ Gaming performance ๐Ÿค– Local AI inference ๐Ÿ–ผ๏ธ Integrated graphics A smaller refresh (Gorgon Halo) is expected in late 2026, with full Medusa Halo likely arriving in 2027. #AMD# #APU# #RDNA5#
Show more
๐Ÿ“ฐ A new coding Agent Index was just released by @ArtificialAnlys. This measures both the model and the harness. No open source harnesses included. OpenSource for Coding is now legit. ๐Ÿ”ฅ Claude Code+GLM-5.1 (53) > Claude Code+Sonnet 4.5 ๐Ÿ”ฅ Claude Code+GLM-5.1 (53) > Gemini CLI+Gemini3.1 DeepSeek V4 Pro & Kimi K2.6 also hit 50
Show more
๐Ÿšจ OMG! Qwen just released Qwen3.6-27B! This is an update to the King ๐Ÿ‘‘ of Open Source Med Models! A 27B dense model that beats the much larger Qwen3.5-397B-A17B (~15x bigger) on coding benchmarks! ๐Ÿคฏ Watch out Opus 4.7 Key wins: โ€ข SWE-Bench Verified: 77.2 (vs 76.2) โ€ข SWE-Bench Pro: 53.5 (vs 50.9) โ€ข Terminal-Bench 2.0: 59.3 (vs 52.5) โœ… Strong multimodal reasoning โœ… Thinking + Non-thinking modes โœ… Apache 2.0 (fully open) Smaller model. Flagship performance. ๐Ÿ”ฅ
Show more