DeepSeek V4 Flash
FeaturedOur most dependable everyday agent: steady on long tool-calling runs, fast, and the cheapest model we serve. It reads text only, so for screenshots and images use DeepSeek V4.1 Flash or GLM 5.3 Flash.
- Price per 1M tokens
- $0.14 in$0.28 out
- $0.028 cache read
- Context
- 1M
- 384K max output
- Parameters
- 284B
- 13B active
- Text only
- Tool calling
- Reasoning optional, default low
How it ranks
Against every model we serve, best first. Starred scores are reported by the publisher on its model card; the rest are independent measurements. About the scores
- GLM 5.345
- Kimi K344
- GLM 5.3 Flash42
- DeepSeek V4.1 Flash39
- DeepSeek V4 Flash34
- Qwen3.6 35B A3B18
- DeepSeek V4.1 Flash74.2%* reported by the publisher
- GLM 5.369%
- Kimi K368.5%
- GLM 5.3 Flash63.4%
- DeepSeek V4 Flash54.4%* reported by the publisher
Not published: Qwen3.6 35B A3B.
- DeepSeek V4.1 Flash90.6%* reported by the publisher
- Kimi K385%
- GLM 5.3 Flash84.3%
- GLM 5.383.9%
- DeepSeek V4 Flash78.7%
- Qwen3.6 35B A3B44.9%
Pricing
- Input
- $0.14 per 1M tokens
- Output
- $0.28 per 1M tokens
- Cache read
- $0.028 per 1M tokens
Specs
- Context window
- 1M tokens
- Max output
- 384K tokens
- Parameters
- 284B total, 13B active per token
- Vision
- No, text only
- Tool calling
- Yes
- Reasoning
- Adjustable (none, low, high, max)
- Default effort
- low
- In production since
- August 3, 2026
Quickstart
Pick your tool; the command already carries this model's id. More tools (Pi, Copilot, Zed, omp) in the setup guides.
OpenCode
Built inLaunch OpenCode on this model with the umans CLI. Umans AI is also built into OpenCode: run /connect and pick it.
curl -fsSL https://api.code.umans.ai/cli/install.sh | bash # once
umans opencode --model umans-deepseek-v4-flash-0731Or consider
- DeepSeek V4.1 Flash$0.15 in · $0.60 out per 1MBest for fast agentic coding at a low price.DeepSWE 74.2%* reported by the publisher
- GLM 5.3$1.40 in · $4.40 out per 1MBest for complex, long-horizon coding.DeepSWE 69%
- Kimi K3$3.00 in · $15.00 out per 1MBest for the hardest, highest-stakes tasks.DeepSWE 68.5%