Register and share your invite link to earn from video plays and referrals.

Inception
@_inception_ai
Pioneering a new generation of LLMs.
10 Following    19K Followers
Today, we’re introducing Mercury 2.5, the most capable diffusion LLM on the market. It offers a 40% jump in intelligence over Mercury 2, and runs over 1,100 tokens/sec on widely-available @NVIDIAAI GPUs. It’s available today on our API, @OpenRouter, and @Baseten. Contact us to evaluate Mercury 2.5 for production:
Show more
Parallel inference is inevitable. @adityagrover_ took the @Ai4Conferences stage to make the case: GPUs parallelized matrix multiplication. Transformers parallelized training. Diffusion parallelizes inference. Sequential token generation has a ceiling. Diffusion breaks it.
Show more