Register and share your invite link to earn from video plays and referrals.

Aliez Ren
@aliez_ren
套利工具 @taoli_tools 池子区间
1.8K Following    16.3K Followers
Introducing GLM-5.2: Frontier Intelligence, Open Weights - Significant improvements in coding and agentic tasks - Strong long-horizon capabilities with a 1M context window - Two levels of reasoning effort: GLM-5.2 (max) pushes the limits, while GLM-5.2 (high) strikes a strong balance between performance and token efficiency - MIT-licensed open weights - Same API pricing as GLM-5.1 Tech Blog: Weights: API: Coding Plan: Chat:
Show more
0
725
13.2K
1.8K
Forward to community
Imagine if codex existed in 1982
Imagine if codex existed in 1998
almost same case but server edition with 11 pcie slots. and 14 fans
Great work! tested on my 4x RTX Pro 6000 (workstation edition but limit power to 300W each) with PCIe 4.0: tp=2, pp=2: prefill 1570, decode 34 tp=4, pp=1: prefill 967, decode 49 my dockerfile:
Show more
GLM-5.1-478B-NVFP4 Running on: - 4x RTX Pro 6000 - Sglang - 370,000 max tokens (1.75x full context) - p10 27.7 | p90 45.6 tok/s decode (gen) - 1340 tok/s prefill I could get 2x decode if I limit to 64k context (100 tok/s) In this video it operates Figma (:
Show more