注册并分享邀请链接,可获得视频播放与邀请奖励。

tonbi
@tonbistudio
No topic is too difficult, it just hasn't been explained well enough yet contributor @ Nous Research Running | Tonbi's AI Garage on YT
加入 October 2021
1.9K 正在关注    19.5K 粉丝
Today’s video is on DFlash 2 from Inco ai and speculative decoding! I explain what DFlash is, how it speeds up local models, and then run an experiment with three versions of a local model, one with no speculative decoding, one with DFlash, and one with DFlash 2 to see the difference! Check it out! @zhijianliu_ Tinkering with DFlash2: How to Speed Up Local AI Models
显示更多