A new open weight Chinese model has hit the timeline!!
This time it’s from Tencent’s Hunyuan team with Hy4 preview.
It’s a massive MoE with 770B total parameters / 49B active per token, a 1M context window, and the weights are released under Apache 2.0.
Important caveat - Tencent says this is still an early Hy4 checkpoint with more pretraining + post training to come.
I actually like how honest their benchmark sheet is. They show plenty of places where the model is still behind despite its size. Hy4 gets 85.4 on Terminal Bench 2.1, basically right in the frontier cluster, and jumps from Hy3’s 28.0 -> 64.3 on DeepSWE. But it’s still behind Kimi K3 at 74.0 and Claude Opus 5 at 74.7 there.
On ProgramBench it gets 17.5 vs Claude’s 39.5, SWE Atlas Refactoring 53.3 vs 60.0, and Humanity’s Last Exam 43.4 vs 53.2.
It’s a 49B active open weight model that is competitive for its size on a bunch of hard coding/agent benchmark.
One thing Tencent also reported on GitHub is they ran a 163 person internal blind eval across 203 engineering tasks where Hy4 slightly beat GLM 5.3 and Kimi K3, which is pretty interesting.
顯示更多