๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

KVCache.AI
@KVCache_AI
Hi, this is official account. We build systems for efficient LLM serving, including KTransformers, Mooncake and AgentENV.
๊ฐ€์ž… August 2018
109 ํŒ”๋กœ์ž‰ ์ค‘    1.1K ํŒฌ
Excited to ship Mooncake alongside Miles v0.1 ๐Ÿš€ Looking forward to pushing large-scale RL infrastructure forward together. Full blog:
Shipping alongside Miles v0.1, Mooncake lands as a new rollout data-transfer backend in Miles, making remote GET 10-14ร— faster than the existing path. With @KVCache_AI, we gave the rollout-to-training handoff a dedicated data plane: - 1.2-1.6ร— faster PUT via structured-object transfer - Structure-aware PUT/GET optimizes serialization and bulk RDMA transfer, including zero-copy reconstruction from registered buffers on GET - Same put/get calls, no change to the RL programming model Read the full blog ๐Ÿ‘‡ link in the comment
๋” ๋ณด๊ธฐ