注册并分享邀请链接,可获得视频播放与邀请奖励。

Tim Dettmers
@Tim_Dettmers
Creator of bitsandbytes. Professor @CarnegieMellon and Research Scientist @allen_ai . I blog about deep learning and PhD life at
加入 October 2012
917 正在关注    48.4K 粉丝
Trying GLM 5.3 right now. Avg: Prefill ~1 ktok/s, thinking/output ~60 tok/s. In our harness, we already see with full thinking traces GLM 5.2 > Fable+Claude Code. But GLM 5.3 is just on another level. It is very precise and concise. Just testing long-task performance. Exciting!
显示更多
Introducing GLM-5.3: Built to Code. Ready for Cyber Defense. - Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model - A major leap in cybersecurity, setting a new standard among open models Tech Blog:
显示更多
0
15
275
18
转发到社区