๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

vLLM
@vllm_project
A high-throughput and memory-efficient inference and serving engine for LLMs. Join to discuss together with the community!
๊ฐ€์ž… March 2024
36 ํŒ”๋กœ์ž‰ ์ค‘    50.3K ํŒฌ
๐Ÿค Day-0 support for MiniCPM5-2B on stable vLLM. โšก Dense 2.6B model with 131K native context ๐Ÿง  Think / No-Think from the same checkpoint ๐Ÿ”ง Tool Calling support via vLLMโ€™s minicpm5 parser Congrats @OpenBMB on the release, and thanks for keeping it on stock LlamaForCausalLM and opening the training data alongside the weights! ๐Ÿ™Œ ๐Ÿ”—
๋” ๋ณด๊ธฐ