“I used my agent to formally verify my software and it fixed all these bugs” is the new “I asked an LLM judge to fix my LLM output”. The devil is in the details and it is insanely difficult to get the formalisms “right” (eg interpretable, expressive, extensible) for humans and agents to co-create software. Seems that we are in for an entirely new age of learning how to build software