Register and share your invite link to earn from video plays and referrals.

Tianwei Yin
@TianweiY
Multimodal AGI. Prev: @MIT @Adobe
156 Following    1.4K Followers
8/ With that, we reframed multimodal generation as structured text/code generation. Diffusion just renders pixels. Planning, logic, reasoning all live in the LLM — so training looks like normal LLM training, and inherits all benefits of it: data + model scaling, reasoning, RL, tool use.
Show more
1/ Our new @reve image model is now #2# on the @arena text-to-image leaderboard — behind only GPT Image 2, ahead of Nano Banana Pro, Microsoft, xAI and everyone else. And it's a 125 point jump over Reve 1.5 from just 3 months ago. The research story behind it 🧵👇
Show more
Today, we’re launching Reve 2.0, the best 4K image model in the world. We invented a new way to generate and edit any image using precise layouts. For the first time, it’s possible to create images you can touch.
Show more