註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Tianwei Yin
@TianweiY
Multimodal AGI. Prev: @MIT @Adobe
加入 January 2019
156 正在關注    1.4K 粉絲
8/ With that, we reframed multimodal generation as structured text/code generation. Diffusion just renders pixels. Planning, logic, reasoning all live in the LLM — so training looks like normal LLM training, and inherits all benefits of it: data + model scaling, reasoning, RL, tool use.
顯示更多