登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Alex Ker 🔭
@thealexker
code+words @baseten | investing in frontiers & sharing my curiosities | prev @bloombergbeta @stanfordhai @neurable.
参加 May 2018
1.3K フォロー中    13.3K ファン
most people forget there are two vectors to optimize for to reduce model cost: 1) reducing the input/output costs, increasing cache hit rates, batching etc 2) packing more intelligence per token (fewer tokens for same task) inference cost = price per token × tokens per task optimizations around the second is underrated and something we’ll only see more of. compression and concision is intelligence.
もっと見る