註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Braintrust
@braintrust
Active observability for agents in production.
加入 August 2023
60 正在關注    7.7K 粉絲
When you eval open-weight models, you also need to eval the serving stack. The inference engine, model revision, precision, caching, and account limits can all impact the system you're testing. We ran an eval on Kimi K3 served on @FireworksAI_HQ vs @Kimi_Moonshot. They performed equally well on quality, but median time to first token was 55% faster on Fireworks. Read more →
顯示更多