註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

vLLM
@vllm_project
A high-throughput and memory-efficient inference and serving engine for LLMs. Join to discuss together with the community!
加入 March 2024
36 正在關注    50.3K 粉絲
🤝 Day-0 support for MiniCPM5-2B on stable vLLM. ⚡ Dense 2.6B model with 131K native context 🧠 Think / No-Think from the same checkpoint 🔧 Tool Calling support via vLLM’s minicpm5 parser Congrats @OpenBMB on the release, and thanks for keeping it on stock LlamaForCausalLM and opening the training data alongside the weights! 🙌 🔗
顯示更多