Register and share your invite link to earn from video plays and referrals.

Zixuan Li
@ZixuanLi_
Lead @Zai_org.
Joined March 2025
240 Following    23.8K Followers
GLM-5.2 delivers a substantial leap in app development capabilities, which also represent demanding long-horizon tasks. Results: - GLM-5.1: 21/70 - GLM-5.2: 48/70 - Claude Fable 5: 56/70 That's more than a twofold improvement from GLM-5.1 to GLM-5.2. These come from an internal benchmark of 35 challenging mobile development tasks, each run twice for a total of 70 trials. We measured task completion, defined as core features working without major issues.
Show more
0
79
1.5K
101
Forward to community