登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Modular
@Modular
Building AI’s unified compute layer. We are hiring → 🚀
参加 January 2022
2 フォロー中    24.3K ファン
Optimizing large scale inference systems is what we do, so we decided to write down what we know. Our LLM Inference Handbook is a free reference covering TTFT, TPOT, goodput, continuous batching, chunked prefill, prefix caching, KV cache math, prefill-decode disaggregation, quantization, and more. It includes 20+ interactive visualizations, is updated continuously, and is open to PRs.
もっと見る