注册并分享邀请链接,可获得视频播放与邀请奖励。

Braintrust
@braintrust
Active observability for agents in production.
加入 August 2023
60 正在关注    7.7K 粉丝
Jev from @typesafeai can replace your LLM-as-a-judge for scoring agent responses. It returns a choice or numeric result with information about uncertainty, so you don't need to spend time and resources prompting a general-purpose model into an LLM judge. Use Jev as a judge scorer in Braintrust and review its selected answer, confidence, and probabilities alongside the score. Trace Jev calls from your own application with the JavaScript or Python SDK. Read more →
显示更多
0
20
185
9
转发到社区