註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

dealign.ai
@dealignai
i like breaking models -
加入 January 2024
64 正在關注    5.5K 粉絲
Prefix caching now stores BY TOOL CALL instead of BY TURN allowing for dramatically speedier prefill speeds. Memory allocator bug has been fixed to 8gb, compiled decode for hybrid and other models enabled by default. Will be focusing heavily in more optimizations where I can in the next few days.
顯示更多