Weโve just released the 1-bit & 4-bit version of Hy3, a flagship-scale 295B model that can be served on a single GPU. ๐
Run Hy3 with llama.cpp, enable MTP, and experience powerful intelligence on dramatically lower hardware.๐๐๐
Canโt wait to see what you build.
#
Hy3# #
Hy# #
GGUF# #
llamacpp#