do you know what you pay for in agentic workloads? cached tokens!
session with 50+ tool calls -> prompt is billed 50 times
all providers give 1/5 cached discount for GLM-5.2
we at
@FireworksAI_HQ dropped it to 1/10, matching GPT/Claude
that's -40% typical savings, have fun!