가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Daniel Han
@danielhanchen
Building @UnslothAI • Making open-source LLMs faster, better & more accessible • YC S24 • ex-NVIDIA ML
가입 April 2016
2K 팔로잉 중    36.3K 팬
Get faster inference with GLM-5.3-Flash GGUFs out of the box in Unsloth Desktop. We enabled MTP and faster long context decoding!
We made GLM-5.3-Flash run 3.3x faster locally! Local GGUF inference is now 1.6–3.4× faster with optimized decoding and bonus multi-token prediction. Run 3-bit on 128GB setups via Unsloth Desktop or llama.cpp. Guide: GGUF:
더 보기