Model selection is a huge driver for AI spend; we just reduced our cost-per-PR from $80 to $30 by switching from Opus 5 to GPT 5.6. We made this decision not by guessing, but by benchmarking. I want to show you how you can do this too!
Join me for a live session on building custom model benchmarks:
Thursday, September 17th 2pm ET
We'll build a custom benchmark that replays your past agent conversations against a set of different models to find the best cost/performance pick. I'll whiteboard the entire setup so you can do this yourself, and we'll walk through the batteries-included version with Warp Factories.
RSVP: 📅