Yesterday, we launched
@GoogleDeepMind's Gemma 4 model on
@cerebras.
The first multimodal model on Cerebras.
1,500 tokens per second.
15x faster than the nearest comparable model.
Multimodal agents can see, reason, act, and retry.
At 1,500 tokens per second, that loop is becoming near instant.
Fast tokens are the most valuable tokens.
And Cerebras delivers the fastest tokens in the world.