This pod was an incredible gift to the community:
not only our first pod about
@xAI, but Ethan really indulged on all our questions on how to train a SOTA Videogen world model, including specific areas (consistent extending/editing, voice) that Grok
@Imagine is *still* SOTA,
on top of the factual overviews he ALSO came loaded with opinions/predictions:
- why he's quitting Videogen for LLMs: video models get most of their intelligence from LLMs, not from scaling video data
- why the next frontier for videogen also happens to be video agent models - agentic models trained to orchestrate video models
- why deterministic compression (like MP4) is a useless target vs VAE compression
- Videomaxxing: if you truly believe in the "Moore's law" of AI/genmedia, then video models become the final boss UI of everything, like Flipbook (below)