注册并分享邀请链接,可获得视频播放与邀请奖励。

Braintrust
@braintrust
The observability layer for production AI.
加入 August 2023
57 正在关注    7.1K 粉丝
We built a Braintrust-native eval in collaboration with Baseten to test whether GLM-5.2 can preserve exact long-context retrieval under production serving constraints. GLM-5.2's retrieval score is effectively flat as context grows from 25K to 50K, which is the result users most want to see from a sparse-attention long-context model. Read the full GLM-5.2 eval →
显示更多