This is true, going open source models is way more silicon intensive, vs usage concentrating with Ant and OpenAI is more efficient, but that efficiency gets captured by the model providers. Some gets passed on to consumers, some gets absorbed by the model players
This is the fundamental tension between Nvidia / Cloud hyperscalers and OAI/Ant
This is the whole on prem to cloud dynamic all over again but on steroids & with much farther reaching implications