Register and share your invite link to earn from video plays and referrals.

Braintrust
@braintrust
The observability layer for production AI.
57 Following    7.1K Followers
We built a Braintrust-native eval in collaboration with Baseten to test whether GLM-5.2 can preserve exact long-context retrieval under production serving constraints. GLM-5.2's retrieval score is effectively flat as context grows from 25K to 50K, which is the result users most want to see from a sparse-attention long-context model. Read the full GLM-5.2 eval →
Show more
Run the best OSS models in Braintrust, in collaboration with @Baseten. Call GLM-5.2 natively, eval its quality, and observe its behavior in production. Save on inference costs without compromising quality by picking the best model for your agent. Free to try through July 31.
Show more