DeepSeek-V4-Flash scores 52 at 11 cents
Merge API reports DeepSeek-V4-Flash scoring 52/100 on the Artificial Analysis Intelligence Index at roughly $0.11 per task, making it the strongest open-weight model below $0.20 per task. DeepSeek positions Flash as its faster, economical model with 284B total parameters, 13B active parameters, and a 1M-token context window.
DeepSeek-V4-Flash’s real advantage is economics: capable enough for serious agent workloads, cheap enough to use at scale. The benchmark result is compelling, but developers should validate reliability and latency on their own task mix before routing production traffic.
- –The Intelligence Index averages nine evaluations spanning agents, coding, general capability, and scientific reasoning, offering broader signal than a single benchmark.
- –Its 284B-parameter architecture with only 13B active at inference helps explain the cost-performance balance.
- –At $0.11 per completed task, Flash becomes especially attractive for high-volume coding agents, automation, and backend workflows.
- –Open weights expand deployment flexibility, but the model’s total size still makes self-hosting a serious infrastructure decision.
- –Cost-per-task remains workload-dependent: long contexts, retries, tool calls, and cache-hit rates can materially change the economics.
DISCOVERED
58m ago
2026-09-11
PUBLISHED
1h ago
2026-09-11
RELEVANCE
AUTHOR
merge_api
