註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Tim Dettmers
@Tim_Dettmers
Creator of bitsandbytes. Professor @CarnegieMellon and Research Scientist @allen_ai . I blog about deep learning and PhD life at
加入 October 2012
917 正在關注    48.4K 粉絲
Trying GLM 5.3 right now. Avg: Prefill ~1 ktok/s, thinking/output ~60 tok/s. In our harness, we already see with full thinking traces GLM 5.2 > Fable+Claude Code. But GLM 5.3 is just on another level. It is very precise and concise. Just testing long-task performance. Exciting!
顯示更多
Introducing GLM-5.3: Built to Code. Ready for Cyber Defense. - Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model - A major leap in cybersecurity, setting a new standard among open models Tech Blog:
顯示更多
0
15
275
18
轉發到社區