deepseek/deepseek-v4-flashDeepSeek-V4-Flash is DeepSeek’s open-weights AI model, released Apr 24, 2026. It costs $0.05/M input and $0.14/M output tokens, and handles a 1M-token context window. Its strongest use case is agentic & tools, where it ranks in the 78th percentile of the models Model Beat tracks.
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window.
Scores on standardized evaluations. Higher is better — and rank shows where DeepSeek-V4-Flash lands among all models tracked on Model Beat.
Graduate-level scientific reasoning
Frontier of human expert knowledge
Common-sense trick questions
Novel ML problem-solving
Scientific research coding
Human-rated web development
Multi-step agentic tasks
Tool-agent-user reliability
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window.
DeepSeek-V4-Flash is an AI model developed by DeepSeek, part of the DeepSeek family, released Apr 24, 2026.
DeepSeek-V4-Flash is served from $0.05 per million input tokens and $0.14 per million output tokens (cheapest credible provider on OpenRouter: OpenInference).
DeepSeek-V4-Flash supports a context window of up to 1M tokens.
On standardized evaluations tracked by Epoch AI, DeepSeek-V4-Flash scores 95.6% on τ²-bench, 35.0% on APEX, 30.3% on Humanity's Last Exam.
DeepSeek-V4-Flash is released with open weights (Open weights (unrestricted)), so the model can be downloaded and self-hosted.
Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Sep 18, 2026.