dotsagent.io
Language:English
DeepSeek · Flagship

DeepSeek-V4.1-Flash API pricing and specs

DeepSeek's newest model, which DeepSeek says beats V4-Pro on quality, speed and price, and which accepts both OpenAI and Anthropic request formats.

Input
⁨$0.30⁩
Cached input
⁨$0.006⁩
Output
⁨$1.20⁩

US dollars per million tokens, standard tier, short-prompt rate

The old names deepseek-v4-flash and deepseek-v4-flash-vision-exp now route here. Cache hits cost $0.006 per million at peak and half that off-peak.

DeepSeek charges half price off-peak. Peak hours are 01:00 to 04:00 and 06:00 to 10:00 UTC on weekdays; every other hour, weekends included, is off-peak. The table shows peak prices.

Spec sheet

API model string
deepseek-flash
Provider
DeepSeek
Context
1M tokens
Max output
384K tokens
Released
September 10, 2026
Batch API
No batch discount listed
Documented agent features
Tool callingStructured outputAdjustable reasoningImage input

What a task costs

Four reference workloads priced at this model's list rates, with caching applied after the first call. Your numbers will differ; the calculator lets you change every assumption.

WorkloadPer taskPer 1,000 tasks
Support bot
4 calls, short tool lookups
⁨$0.007⁩⁨$7⁩
Coding agent
30 calls, files and test output
⁨$0.218⁩⁨$218⁩
Browser agent
25 calls, screenshots and page text
⁨$0.186⁩⁨$186⁩
Research agent
15 calls, long search results
⁨$0.180⁩⁨$180⁩

Change the assumptions

Other DeepSeek models

Similar cost elsewhere

Models from other providers whose cost on the 30-step coding workload is closest to this one.

Sources

  1. api-docs.deepseek.com/quick_start/pricing
  2. api-docs.deepseek.com/updates
  3. api-docs.deepseek.com/news/news260910

Independent reference for people who build AI agents. Not affiliated with any vendor named here.

© 2026 DotsAgent · Facts checked October 1, 2026