rare is the team doing true architectural AI research these days. I talk to the extraordinarily broad and productive researcher Stefano Ermon about his company, latency, and why it’s still worth going after good ideas in the age of scaling
Today's LLMs can't generate token 10 until token 9 exists.
@_inception_ai CEO @StefanoErmon thinks he's found the winner: diffusion LLMs that generate tokens in parallel.
New episode of No Priors, out now.