Deepseek V4 Flash 0731 now has the best model intelligence vs cost out of any model.
It’s around the same intelligence as GLM 5.2 and GPT Luna while being way cheaper (only $0.14/$0.28 per 1M tokens).
Absolutely insane release.
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!
🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇
🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex!
Check out the configuration details in our official API docs: