GPT-6 Astra API pricing and specs
OpenAI's most capable API model, with the full Responses API toolset: hosted shell, apply patch, skills, tool search, web and file search, computer use and remote MCP.
Cache writes cost $12.50 per million. A separate Ultrafast tier runs at $60 input and $300 output, Fast mode doubles the price, and EU or other regional processing adds 10%. Astra has no 'none' reasoning setting and ignores custom temperature.
Prompts over 272K tokens are billed at $20 input, $2 cached and $75 output per million, and the higher rate applies to the whole request.
Spec sheet
- API model string
gpt-6-astra- Provider
- OpenAI
- Context
- 1.05M tokens
- Max output
- 128K tokens
- Released
- September 8, 2026
- Batch API
- 50% off, results within hours
- Documented agent features
- Tool callingStructured outputComputer useRemote MCPBuilt-in web searchAdjustable reasoning
What a task costs
Four reference workloads priced at this model's list rates, with caching applied after the first call. Your numbers will differ; the calculator lets you change every assumption.
| Workload | Per task | Per 1,000 tasks |
|---|---|---|
| Support bot | $0.280 | $280 |
| Coding agent | $9.89 | $9,887 |
| Browser agent | $6.98 | $6,983 |
| Research agent | $6.62 | $6,615 |
Other OpenAI models
Similar cost elsewhere
Models from other providers whose cost on the 30-step coding workload is closest to this one.