AI video is becoming another compute story.
Text generation was only the first wave.
Images, video, voice, and real-time multimodal models require much heavier inference.
The more AI moves from typing answers to generating media and actions,
the more compute demand expands.