Ling-3.0-flash is live on DeepInfra
A 124B MoE from
@AntLingAGI with just 5.1B active params, large-model capacity at close to small-model cost. Reasoning + tool calling on by default, 128K native context (up to 1M with YaRN).
Built for agents. Live now 👇