๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Red Hat AI
@RedHat_AI
Accelerating AI innovation with open platforms and community. The future of AI is open.
๊ฐ€์ž… May 2018
2.1K ํŒ”๋กœ์ž‰ ์ค‘    11.7K ํŒฌ
Red Hat AI just shipped DFlash speculator checkpoints for two of @NVIDIAAI's most powerful open models: โ†’ Nemotron Ultra 550B โ†’ Nemotron Super 120B On math and reasoning: ~5 out of 7 draft tokens accepted on average. On code (HumanEval): ~3.4 out of 7. Both checkpoints trained with the open source Speculators library from @vllm_project. Apache 2.0. Validated on NVIDIA B200. One flag to enable in vLLM: --spec-model RedHatAI/NVIDIA-Nemotron-3-Ultra-550B-A55B-speculator.dflash --spec-tokens 7 --spec-method dflash ๐Ÿ”— Ultra 550B: ๐Ÿ”— Super 120B:
๋” ๋ณด๊ธฐ