注册并分享邀请链接,可获得视频播放与邀请奖励。

Unsloth AI
@UnslothAI
Train and run models locally! 🦥
加入 November 2023
480 正在关注    88.9K 粉丝
Google releases Gemma 4 QAT. ✨ You can now run Gemma 4 at 3x less memory with near original performance. Quantization-Aware Training (QAT) makes it possible to run Gemma 4 26B-A4B on 16GB RAM. GGUFs: QAT Guide:
显示更多
We just dropped Gemma 4 Quantization-Aware Training (QAT) checkpoints on Hugging Face! All Gemma 4 model sizes and their drafters are now optimized with QAT to cut memory requirements and maximize on-device performance!
显示更多
0
93
2.9K
411
转发到社区