dotsagent.io
Language:English
Google · Workhorse

Gemini 3.5 Flash API pricing and specs

The Flash model that Gemini 3.8 Flash replaced at the top, now listed by Google as legacy.

Input
⁨$1.50⁩
Cached input
⁨$0.15⁩
Output
⁨$9⁩

US dollars per million tokens, standard tier, short-prompt rate

Until the end of 2026 it costs more than Gemini 3.8 Flash, which is on an introductory price, so there is little reason to start a new project on it.

Spec sheet

API model string
gemini-3.5-flash
Provider
Google
Context
1.05M tokens
Max output
66K tokens
Released
May 19, 2026
Batch API
50% off, results within hours
Documented agent features
Tool callingStructured outputComputer useBuilt-in web searchAdjustable reasoning

What a task costs

Four reference workloads priced at this model's list rates, with caching applied after the first call. Your numbers will differ; the calculator lets you change every assumption.

WorkloadPer taskPer 1,000 tasks
Support bot
4 calls, short tool lookups
⁨$0.046⁩⁨$46⁩
Coding agent
30 calls, files and test output
⁨$1.54⁩⁨$1,537⁩
Browser agent
25 calls, screenshots and page text
⁨$1.06⁩⁨$1,062⁩
Research agent
15 calls, long search results
⁨$1.03⁩⁨$1,026⁩

Change the assumptions

Other Google models

Similar cost elsewhere

Models from other providers whose cost on the 30-step coding workload is closest to this one.

Sources

  1. ai.google.dev/gemini-api/docs/pricing
  2. ai.google.dev/gemini-api/docs/models/gemini-3.5-flash
  3. blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-5/

Independent reference for people who build AI agents. Not affiliated with any vendor named here.

© 2026 DotsAgent · Facts checked October 1, 2026