注册并分享邀请链接,可获得视频播放与邀请奖励。

jietang
@jietang
Professor @ Tsinghua, Founder of AGI, LLM. “The value of a man should be seen in what he gives and not in what he is able to receive.”―Einstein
加入 May 2008
394 正在关注    56.1K 粉丝
5.2 could be better with more RL ...
Deepswe's benchmark results are my own experience. I've used all models, GLM 5.2 ≈ Claude Opus 4.6–4.7. Kimi 2.7 code more like inference optimization. Looking forward to K3. Doubao-seed 2.1 Pro around 37% ≈ Gemini 3.5 Flash. code are quite weak, but visual are strong.
显示更多
0
79
1.4K
63
转发到社区