DeepSeek V4 Flash

Featured

Open weights by DeepSeek

Our most dependable everyday agent: steady on long tool-calling runs, fast, and the cheapest model we serve. It reads text only, so for screenshots and images use DeepSeek V4.1 Flash or GLM 5.3 Flash.

Price per 1M tokens
$0.14 in$0.28 out
$0.028 cache read
Context
1M
384K max output
Parameters
284B
13B active
Weights
FP4 + FP8
Official release
  • Text only
  • Tool calling
  • Reasoning optional, default low

How it ranks

Against every model we serve, best first. Starred scores are reported by the publisher on its model card; the rest are independent measurements. About the scores

Artificial Analysis Intelligence Indexby Artificial Analysis
  1. GLM 5.345
  2. Kimi K344
  3. GLM 5.3 Flash42
  4. DeepSeek V4.1 Flash39
  5. DeepSeek V4 Flash34
  6. Qwen3.6 35B A3B18
DeepSWEby Datacurve
  1. DeepSeek V4.1 Flash74.2%* reported by the publisher
  2. GLM 5.369%
  3. Kimi K368.5%
  4. GLM 5.3 Flash63.4%
  5. DeepSeek V4 Flash54.4%* reported by the publisher

Not published: Qwen3.6 35B A3B.

Terminal-Bench 2.1by tbench.ai
  1. DeepSeek V4.1 Flash90.6%* reported by the publisher
  2. Kimi K385%
  3. GLM 5.3 Flash84.3%
  4. GLM 5.383.9%
  5. DeepSeek V4 Flash78.7%
  6. Qwen3.6 35B A3B44.9%
  • Artificial Analysis Intelligence Index for DeepSeek V4 Flash: independent measurement. Source
  • DeepSWE for DeepSeek V4 Flash: reported by the publisher on its model card. Source
  • Terminal-Bench 2.1 for DeepSeek V4 Flash: independent measurement. Source

Pricing

Input
$0.14 per 1M tokens
Output
$0.28 per 1M tokens
Cache read
$0.028 per 1M tokens

See all prices

Specs

Context window
1M tokens
Max output
384K tokens
Parameters
284B total, 13B active per token
Vision
No, text only
Tool calling
Yes
Reasoning
Adjustable (none, low, high, max)
Default effort
low
In production since
August 3, 2026

Quickstart

Pick your tool; the command already carries this model's id. More tools (Pi, Copilot, Zed, omp) in the setup guides.

OpenCode
Built in
Launch OpenCode on this model with the umans CLI. Umans AI is also built into OpenCode: run /connect and pick it.
Terminal
curl -fsSL https://api.code.umans.ai/cli/install.sh | bash   # once
umans opencode --model umans-deepseek-v4-flash-0731

Or consider

DeepSeek V4 Flash API: pricing, context and benchmarks · Umans AI