Programmatic tool calling (aka code mode) is here for OpenAI models via the Responses API.
We've seen significant gains in token efficiency (lower cost and latency) when using it and because it runs in in-memory V8s, it's highly performant and ZDR-compatible with no additional container costs.
Programmatic Tool Calling lets GPT-5.6 write and run JavaScript to coordinate complex tool workflows, process results, and decide what to do next.
That means fewer model round trips, fewer tokens, and less orchestration for developers.
Programs run in isolated hosted V8 runtimes, so developers get stateless code-driven orchestration without managing containers.