Register and share your invite link to earn from video plays and referrals.

Baseten
@baseten
Inference is everything.
Joined March 2021
81 Following    17.2K Followers
Some workloads demand either the highest throughput or the lowest latency. Embedding workloads need both. We built Baseten Embeddings Inference (BEI) to meet that need, and we're thrilled to partner with @turbopuffer to power BEI-optimized models in tpuf!
Show more
now in beta: native embeddings in tpuf embedding is the most painful part of puffing. we want to make it easy you can now convert chunks to vectors as you read and write to turbopuffer, without extra calls to an embedding model provider API docs:
Show more