ModelBeat
Models/Ling-3.0-flash
All models
L
inclusionai/ling-3.0-flash

Ling-3.0-flash is an AI model, released Jul 23, 2026. It costs $0.06/M input and $0.18/M output tokens, and handles a 262K-token context window.

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.

Context
262K
Input
$0.06 / 1M
Output
$0.18 / 1M
Released
Jul 23, 2026
Modalities
texttext

Benchmarks

Scores on standardized evaluations. Higher is better — and rank shows where Ling-3.0-flash lands among all models tracked on Model Beat.

39.0
Intelligence Index
Artificial Analysis
39th percentile of tracked models
35.0
Coding Index
Artificial Analysis
35th percentile of tracked models

Reasoning

2 evals
GPQA Diamond85.5%

Graduate-level scientific reasoning

Humanity's Last Exam23.7%

Frontier of human expert knowledge

Coding

1 evals
SciCode42.0%

Scientific research coding

In the news

1

Changelog

6
  • Ling-3.0-flash: cheapest credible provider via DeepInfra now $0.06 per 1M input tokens (was $0.075, -20%)cheapest provider · Aug 11, 2026
  • Ling-3.0-flash: cheapest credible provider via DeepInfra now $0.18 per 1M output tokens (was $0.22, -18%)cheapest provider · Aug 11, 2026
  • Ling-3.0-flash: Humanity's Last Exam improved from 22.1% to 23.7% (+7%)benchmark · Aug 6, 2026
  • Ling-3.0-flash: context window changed from 131,072 to 262,144 tokenscontext · Aug 6, 2026
  • Ling-3.0-flash: context window changed from 262,144 to 131,072 tokenscontext · Aug 6, 2026
  • Ling-3.0-flash added to the tracker: 262K context.tracked · Jul 23, 2026

Compare Ling-3.0-flash with

Frequently asked questions

What is Ling-3.0-flash?

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.

How much does Ling-3.0-flash cost?

Ling-3.0-flash is served from $0.06 per million input tokens and $0.18 per million output tokens (cheapest credible provider on OpenRouter: DeepInfra).

What is Ling-3.0-flash's context window?

Ling-3.0-flash supports a context window of up to 262K tokens.

How does Ling-3.0-flash perform on benchmarks?

On standardized evaluations tracked by Epoch AI, Ling-3.0-flash scores 85.5% on GPQA Diamond, 23.7% on Humanity's Last Exam, 42.0% on SciCode.

Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Sep 10, 2026.