登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Xichen Pan
@xichen_pan
PhD Student @NYU_Courant, Researcher @AIatMeta; Multimodal Generation | Prev: @MSFTResearch, @AlibabaGroup, @sjtu1896; More at
参加 August 2022
584 フォロー中    746 ファン
Modern text-to-image models are increasingly powered by large pretrained LLMs. But there is a curious mismatch: the LLM typically encodes the prompt only once, while the evolving noisy latent states are handled entirely by a newly trained generative backbone. Can pretrained multimodal prior participate in the denoising process? Introducing RepFusion. (1/12) 📄 🌐
もっと見る