註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

(((ل()(ل() 'yoav))))👾
@yoavgo
加入 May 2009
2.2K 正在關注    89.6K 粉絲
why would cache reads be cheaper (assuming its not only business decision)? smaller cache? distilled model? shallower model? fewer tokens?
Cache reads with Fable 5.1 cost 75% less than Fable 5’s. This reduces the cost of the model in practice by around 25% for typical workloads, and up to 45% for highly agentic ones.
顯示更多
0
10
15
1
轉發到社區