註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Baseten
@baseten
Inference is everything.
加入 March 2021
81 正在關注    17.2K 粉絲
Some workloads demand either the highest throughput or the lowest latency. Embedding workloads need both. We built Baseten Embeddings Inference (BEI) to meet that need, and we're thrilled to partner with @turbopuffer to power BEI-optimized models in tpuf!
顯示更多
now in beta: native embeddings in tpuf embedding is the most painful part of puffing. we want to make it easy you can now convert chunks to vectors as you read and write to turbopuffer, without extra calls to an embedding model provider API docs:
顯示更多