注册并分享邀请链接,可获得视频播放与邀请奖励。

Pika
@pika_labs
Tools for creation that happen to be AI
加入 April 2023
103 正在关注    151K 粉丝
Under the hood: Text → transformer prompt conditioning Generation → text-conditioned diffusion transformer in compressed acoustic latent space Output → semantic-acoustic autoencoder decodes a 44.1 kHz stereo waveform Flow matching, teacher–student distillation, and post-training with human feedback reduce the generation path to a few high-quality steps.
显示更多