注册并分享邀请链接,可获得视频播放与邀请奖励。

(((ل()(ل() 'yoav))))👾
@yoavgo
加入 May 2009
2.2K 正在关注    89.6K 粉丝
why would cache reads be cheaper (assuming its not only business decision)? smaller cache? distilled model? shallower model? fewer tokens?
Cache reads with Fable 5.1 cost 75% less than Fable 5’s. This reduces the cost of the model in practice by around 25% for typical workloads, and up to 45% for highly agentic ones.
显示更多
0
10
15
1
转发到社区