登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

vLLM
@vllm_project
A high-throughput and memory-efficient inference and serving engine for LLMs. Join to discuss together with the community!
参加 March 2024
36 フォロー中    50.4K ファン
⚡ New RTX local-agent optimizations include a vLLM speedup on Blackwell. @NVIDIARTXSpark reports: 🛠️ 1.2x vLLM performance on RTX PRO 6000 Blackwell ⚡ Up to 1.4x on a two-system DGX Spark cluster Great to see the local serving path getting faster across RTX and DGX Spark.
もっと見る