---
title: "GPT-6 Luna API pricing: ⁨$0.10⁩ in, ⁨$0.50⁩ out · DotsAgent"
description: "GPT-6 Luna by OpenAI: price per million input, cached and output tokens, context window, agent features and what a real agent task costs on it."
url: https://dotsagent.io/models/gpt-6-luna
---

[OpenAI](https://dotsagent.io/providers/openai)

· Small and fast

# GPT-6 Luna API pricing and specs

OpenAI's small model, with the same hosted tools as Astra and a 'none' reasoning setting for cheap, fast steps.

Input

$0.10

Cached input

$0.010

Output

$0.50

US dollars per million tokens, standard tier, short-prompt rate

In Chat Completions, function calling works only with reasoning effort set to none; use the Responses API to combine tools with reasoning. Knowledge cutoff is 18 May 2026.

Prompts over 272K tokens are billed at $0.20 input, $0.020 cached and $0.75 output per million, and the higher rate applies to the whole request.

## Spec sheet

- **API model string**: `gpt-6-luna`
- **Provider**: [OpenAI](https://dotsagent.io/providers/openai)
- **Context**: 1.05M tokens
- **Max output**: 128K tokens
- **Released**: September 22, 2026
- **Batch API**: 50% off, results within hours
- **Documented agent features**: Tool callingStructured outputComputer useRemote MCPBuilt-in web searchAdjustable reasoning

## What a task costs

Four reference workloads priced at this model's list rates, with caching applied after the first call. Your numbers will differ; the calculator lets you change every assumption.

| Workload | Per task | Per 1,000 tasks |
| --- | --- | --- |
| Support bot 4 calls, short tool lookups | $0.003 | $3 |
| Coding agent 30 calls, files and test output | $0.099 | $99 |
| Browser agent 25 calls, screenshots and page text | $0.070 | $70 |
| Research agent 15 calls, long search results | $0.066 | $66 |

[Change the assumptions](https://dotsagent.io/cost#model=gpt-6-luna)

## Other OpenAI models

- [GPT-6 Astra](https://dotsagent.io/models/gpt-6-astra): $10 / $50

- [GPT-6.1 Sol](https://dotsagent.io/models/gpt-6-1-sol): $2 / $10

## Similar cost elsewhere

Models from other providers whose cost on the 30-step coding workload is closest to this one.

- [Qwen3.7-Flash](https://dotsagent.io/models/qwen3-7-flash): $0.101 per coding task

- [GLM-5.3-Flash](https://dotsagent.io/models/glm-5-3-flash): $0.182 per coding task

- [DeepSeek-V4.1-Flash](https://dotsagent.io/models/deepseek-v4-1-flash): $0.218 per coding task

- [Gemini 3.5 Flash-Lite](https://dotsagent.io/models/gemini-3-5-flash-lite): $0.333 per coding task

## Sources

1. [developers.openai.com](https://developers.openai.com/api/docs/models/gpt-6-luna)/api/docs/models/gpt-6-luna
2. [developers.openai.com](https://developers.openai.com/api/docs/pricing)/api/docs/pricing
3. [developers.openai.com](https://developers.openai.com/api/docs/changelog)/api/docs/changelog

Independent reference for people who build AI agents. Not affiliated with any vendor named here.

© 2026 DotsAgent · Facts checked October 1, 2026
