ModelBeat
Models/DeepSeek-V4-Flash
All models
D

DeepSeek: DeepSeek-V4-Flash

deepseek/deepseek-v4-flash

DeepSeek-V4-Flash is DeepSeek’s open-weights AI model, released Apr 24, 2026. It costs $0.14/M input and $0.28/M output tokens, and handles a 1M-token context window. Its strongest use case is agentic & tools, where it ranks in the 81st percentile of the models Model Beat tracks.

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window.

Context
1M
Input
$0.14 / 1M
Output
$0.28 / 1M
Cheapest via SiliconFlow
$0.13 / $0.28 / 1M
Released
Apr 24, 2026
Modalities
texttext
Strongest atAgentic & Tools top 19%Coding top 40%

Benchmarks

Scores on standardized evaluations. Higher is better — and rank shows where DeepSeek-V4-Flash lands among all models tracked on Model Beat.

66.0
Intelligence Index
Epoch AI
66th percentile of tracked models
60.0
Coding Index
Epoch AI
60th percentile of tracked models
81.0
Agentic Index
Epoch AI
81st percentile of tracked models

Reasoning

3 evals
GPQA Diamond89.4%

Graduate-level scientific reasoning

Humanity's Last Exam32.1%

Frontier of human expert knowledge

WeirdML45.6%

Novel ML problem-solving

Coding

1 evals
SciCode44.9%

Scientific research coding

Agentic & Tools

1 evals
τ²-bench95.0%

Tool-agent-user reliability

Changelog

6
  • DeepSeek-V4-Flash: cheapest credible provider via SiliconFlow now $0.13 per 1M input tokens (was $0.0966, +35%)cheapest provider · Jul 26, 2026
  • DeepSeek-V4-Flash: cheapest credible provider via SiliconFlow now $0.28 per 1M output tokens (was $0.1932, +45%)cheapest provider · Jul 26, 2026
  • DeepSeek-V4-Flash: cheapest credible provider via StreamLake now $0.0966 per 1M input tokens (was $0.077, +25%)cheapest provider · Jul 15, 2026
  • DeepSeek-V4-Flash: cheapest credible provider via StreamLake now $0.1932 per 1M output tokens (was $0.154, +25%)cheapest provider · Jul 15, 2026
  • DeepSeek-V4-Flash: cheapest credible provider via StreamLake now $0.077 per 1M input tokens (was $0.098, -21%)cheapest provider · Jul 13, 2026
  • DeepSeek-V4-Flash: cheapest credible provider via StreamLake now $0.154 per 1M output tokens (was $0.196, -21%)cheapest provider · Jul 13, 2026

Compare DeepSeek-V4-Flash with

Frequently asked questions

What is DeepSeek-V4-Flash?

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window.

Who created DeepSeek-V4-Flash?

DeepSeek-V4-Flash is an AI model developed by DeepSeek, part of the DeepSeek family, released Apr 24, 2026.

How much does DeepSeek-V4-Flash cost?

DeepSeek-V4-Flash's list price from DeepSeek is $0.14 per million input tokens and $0.28 per million output tokens. The cheapest credible third-party provider on OpenRouter (SiliconFlow) serves it at $0.13/$0.28 per 1M.

What is DeepSeek-V4-Flash's context window?

DeepSeek-V4-Flash supports a context window of up to 1M tokens.

How does DeepSeek-V4-Flash perform on benchmarks?

On standardized evaluations tracked by Epoch AI, DeepSeek-V4-Flash scores 95.0% on τ²-bench, 89.4% on GPQA Diamond, 32.1% on Humanity's Last Exam.

Is DeepSeek-V4-Flash open source?

DeepSeek-V4-Flash is released with open weights (Open weights (unrestricted)), so the model can be downloaded and self-hosted.

Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Jul 30, 2026.