Register and share your invite link to earn from video plays and referrals.

Tianwei Yin
@TianweiY
Multimodal AGI. Prev: @MIT @Adobe
Joined January 2019
156 Following    1.4K Followers
8/ With that, we reframed multimodal generation as structured text/code generation. Diffusion just renders pixels. Planning, logic, reasoning all live in the LLM — so training looks like normal LLM training, and inherits all benefits of it: data + model scaling, reasoning, RL, tool use.
Show more