註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

tonbi
@tonbistudio
No topic is too difficult, it just hasn't been explained well enough yet contributor @ Nous Research Running | Tonbi's AI Garage on YT
加入 October 2021
1.9K 正在關注    19.5K 粉絲
Today’s video is on DFlash 2 from Inco ai and speculative decoding! I explain what DFlash is, how it speeds up local models, and then run an experiment with three versions of a local model, one with no speculative decoding, one with DFlash, and one with DFlash 2 to see the difference! Check it out! @zhijianliu_ Tinkering with DFlash2: How to Speed Up Local AI Models
顯示更多