註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Pika
@pika_labs
Tools for creation that happen to be AI
加入 April 2023
103 正在關注    151K 粉絲
Under the hood: Text → transformer prompt conditioning Generation → text-conditioned diffusion transformer in compressed acoustic latent space Output → semantic-acoustic autoencoder decodes a 44.1 kHz stereo waveform Flow matching, teacher–student distillation, and post-training with human feedback reduce the generation path to a few high-quality steps.
顯示更多