dotsagent.io
Language:English
Google · Flagship

Gemini 3.8 Flash API pricing and specs

Google's most capable model in the public Gemini API, with function calling, computer use in preview, code execution and grounding in Search and Maps.

Input
⁨$0.75⁩
Cached input
⁨$0.075⁩
Output
⁨$3.75⁩

US dollars per million tokens, standard tier, short-prompt rate

Output price includes thinking tokens. Search grounding gives 5,000 free queries a month across Gemini 3.x models, then costs $14 per thousand. The 'minimal' thinking level returns an error.

This is an introductory price. From January 1, 2027 the list price becomes ⁨$1.50⁩ input, ⁨$0.15⁩ cached and ⁨$7.50⁩ output per million tokens.

Spec sheet

API model string
gemini-3.8-flash
Provider
Google
Context
1.05M tokens
Max output
66K tokens
Released
September 2, 2026
Batch API
50% off, results within hours
Documented agent features
Tool callingStructured outputComputer useBuilt-in web searchAdjustable reasoning

What a task costs

Four reference workloads priced at this model's list rates, with caching applied after the first call. Your numbers will differ; the calculator lets you change every assumption.

WorkloadPer taskPer 1,000 tasks
Support bot
4 calls, short tool lookups
⁨$0.021⁩⁨$21⁩
Coding agent
30 calls, files and test output
⁨$0.742⁩⁨$742⁩
Browser agent
25 calls, screenshots and page text
⁨$0.524⁩⁨$524⁩
Research agent
15 calls, long search results
⁨$0.496⁩⁨$496⁩

Change the assumptions

Other Google models

Similar cost elsewhere

Models from other providers whose cost on the 30-step coding workload is closest to this one.

Sources

  1. ai.google.dev/gemini-api/docs/pricing
  2. ai.google.dev/gemini-api/docs/models/gemini-3.8-flash
  3. ai.google.dev/gemini-api/docs/changelog

Independent reference for people who build AI agents. Not affiliated with any vendor named here.

© 2026 DotsAgent · Facts checked October 1, 2026