It's the summer of Flash models.
First GLM, now DeepSeek, these midsize models push the Pareto Frontier with remarkable intelligence at low prices.
And, they outperform full-size last-gen models w/ new efficient architectures. Excited to see these recipes scale to 2T+ params.
🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient.
🔹 Introducing the smallest model in our new architecture family, with native visual understanding.
🔹 Designed for greater capability, faster inference, higher throughput, and scaling to larger models.
1/6