inclusionai/ling-3.0-flashLing-3.0-flash is an AI model, released Jul 23, 2026. It costs $0.06/M input and $0.18/M output tokens, and handles a 262K-token context window.
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.
Scores on standardized evaluations. Higher is better — and rank shows where Ling-3.0-flash lands among all models tracked on Model Beat.
Graduate-level scientific reasoning
Frontier of human expert knowledge
Scientific research coding
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.
Ling-3.0-flash is served from $0.06 per million input tokens and $0.18 per million output tokens (cheapest credible provider on OpenRouter: DeepInfra).
Ling-3.0-flash supports a context window of up to 262K tokens.
On standardized evaluations tracked by Epoch AI, Ling-3.0-flash scores 85.5% on GPQA Diamond, 23.7% on Humanity's Last Exam, 42.0% on SciCode.
Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Sep 10, 2026.