DeepSeek · Flagship
DeepSeek-V4.1-Flash API pricing and specs
DeepSeek's newest model, which DeepSeek says beats V4-Pro on quality, speed and price, and which accepts both OpenAI and Anthropic request formats.
$0.30
$0.006
$1.20
The old names deepseek-v4-flash and deepseek-v4-flash-vision-exp now route here. Cache hits cost $0.006 per million at peak and half that off-peak.
DeepSeek charges half price off-peak. Peak hours are 01:00 to 04:00 and 06:00 to 10:00 UTC on weekdays; every other hour, weekends included, is off-peak. The table shows peak prices.
Spec sheet
- API model string
deepseek-flash- Provider
- DeepSeek
- Context
- 1M tokens
- Max output
- 384K tokens
- Released
- September 10, 2026
- Batch API
- No batch discount listed
- Documented agent features
- Tool callingStructured outputAdjustable reasoningImage input
What a task costs
Four reference workloads priced at this model's list rates, with caching applied after the first call. Your numbers will differ; the calculator lets you change every assumption.
| Workload | Per task | Per 1,000 tasks |
|---|---|---|
| Support bot | $0.007 | $7 |
| Coding agent | $0.218 | $218 |
| Browser agent | $0.186 | $186 |
| Research agent | $0.180 | $180 |
Other DeepSeek models
Similar cost elsewhere
Models from other providers whose cost on the 30-step coding workload is closest to this one.