注册并分享邀请链接,可获得视频播放与邀请奖励。

NVIDIA AI
@NVIDIAAI
Teaching your AI new tricks.
加入 June 2016
898 正在关注    343.5K 粉丝
How can a 30B-parameter model activate just 3B parameters per token and still draw on the full model’s capacity? Learn how dense and MoE models use parameters differently, and what that means for throughput, memory and serving complexity. Check out our new technical explainer:
显示更多