註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Lysandre
@LysandreJik
Chief Open-Source Officer (COSO) at Hugging Face
加入 March 2019
647 正在關注    12.5K 粉絲
Transformers has supported loading GGUF files for a few years now, by unquantizing them. Thanks to @_marcsun, we're now using GGML kernels through the `kernels` library to run at the same performance as llama.cpp Huge kudos to the entire @ggml_org for making these kernels!
顯示更多
0
8
51
17
轉發到社區