Another thought I had yesterday in that conversation:
These models are so much more now than text-to-code converters.
Think bigger! Aim higher!
Now that the models can contribute across your whole stack, vertically, it's to let them loose horizontally:
Let them monitor deployments, let them debug prod, let them do ops, let them help end to end.
If you think "they'll screw this up" that's on you and your codebase. Either you need to spend more tokens and build up instincts or you need to make your codebase and company processes friendlier to agents.
Unexpectedly found myself in a conversation about agents yesterday evening with someone who now has to use Qwen at work, because company wants to save money.
Really struggled with expressing how much of a category change using latest frontier models is vs. using Qwen.