GPT-6 pricing: Astra, Sol and Luna

Updated September 23, 2026 · first published September 23, 2026

OpenAI's GPT-6 family now has three models. GPT-6 Astra launched on 3 September 2026. GPT-6 Sol and GPT-6 Luna followed on 22 September 2026, about 90 minutes after Anthropic released Claude Opus 5.5. This page lists every GPT-6 API price from OpenAI's pricing page, read on 23 September 2026, and shows where the bill actually comes from.

GPT-6 API prices per 1M tokens

Standard tier, prompts up to 272K input tokens.

ModelInputCached inputCache writeOutput
GPT-6 Astra$10.00$1.00$12.50$50.00
GPT-6 Sol$2.00$0.20$2.50$10.00
GPT-6 Luna$0.10$0.01$0.125$0.50

The pattern is the same on all three: cached input is 10% of input, a cache write is 1.25× input, and output is 5× input. Astra costs 5× Sol, and Sol costs 20× Luna.

GPT-6 vs GPT-5.6: what changed

TierGPT-5.6 input / outputGPT-6 input / outputChange
Sol$4.00 / $20.00$2.00 / $10.00−50% on both
Luna$0.20 / $1.20$0.10 / $0.50−50% input, −58% output
Terra$2.00 / $12.00no GPT-6 Terra

OpenAI's pricing page notes that the GPT-5.6 Sol price above is promotional and runs at least through 21 November 2026. GPT-5.6 Terra has no GPT-6 successor yet. GPT-6 Sol now costs the same as Terra on input and less on output, so Terra workloads are the first candidates to move.

Batch, flex and fast mode

ModelBatch or flex (in / cached / out)Fast mode (in / cached / out)
GPT-6 Astra$5.00 / $0.50 / $25.00$20.00 / $2.00 / $100.00
GPT-6 Sol$1.00 / $0.10 / $5.00$4.00 / $0.40 / $20.00
GPT-6 Luna$0.05 / $0.005 / $0.25$0.20 / $0.02 / $1.00

Batch and flex are half price. Fast mode is double. Fast mode is the old priority processing, renamed on 30 July 2026, and both service_tier: "priority" and service_tier: "fast" still work. That rename matters for cost tracking: a dashboard filtering on the old tier name can miss fast-mode spend. Data residency endpoints add 10% for models released on or after 5 March 2026, which includes all of GPT-6.

The 272K long-context price jump

All three models have a 1,050,000-token context window, a 922,000-token input limit and 128,000 max output tokens. Above 272K input tokens, the whole request is billed at long-context rates: 2× input and cached input, 1.5× output.

ModelLong-context inputCachedOutput
GPT-6 Astra$20.00$2.00$75.00
GPT-6 Sol$4.00$0.40$15.00
GPT-6 Luna$0.20$0.02$0.75

This is a step, not a slope. Take one GPT-6 Sol request with 90% of the prompt cached and 20K output tokens. At 270K input it costs $0.303. At 300K input it costs $0.528. Eleven percent more context costs 74% more. The same 74% jump applies to Astra and Luna, because the multipliers are the same. Agents that let context grow unchecked cross 272K without anyone deciding to. Trim or summarise history before that line, and alert on requests above it.

Specs that affect cost

ModelKnowledge cutoffReasoning effortTier 1 limit
GPT-6 Astra30 April 2026low, medium, high, xhigh, max500 RPM, 500K TPM
GPT-6 Sol20 April 2026none, low, medium (default), high, xhigh, max500 RPM, 500K TPM
GPT-6 Luna18 May 2026none, low, medium (default), high, xhigh, max500 RPM, 500K TPM

Astra has no none setting, so it always reasons. The effort setting moves the bill as much as the model does: on the Artificial Analysis Intelligence Index, GPT-6 Sol costs $0.13 per task at low effort and $1.06 at max. See real cost per task for every setting.

What one agent step costs

One step with 200K input tokens, 90% cached, and 20K output tokens, at standard list prices.

ModelCost per stepPer 10,000 steps
GPT-6 Luna$0.014$138
GPT-6 Sol$0.276$2,760
Claude Opus 5.5$0.516$5,160
GPT-5.6 Sol$0.552$5,520
GPT-6 Astra$1.380$13,800

ChatGPT and Codex are billed separately

GPT-6 Sol and Luna are rolling out in ChatGPT Work and Codex for Plus, Pro, Business, Enterprise and Edu users, and Free and Go users get Luna in the desktop app. Those are subscription plans. None of the prices on this page apply to them, and API usage is never covered by a ChatGPT plan.

Which GPT-6 model to use

Prices change. The LLM API pricing tracker reads OpenAI's pricing page live and picks up new GPT-6 models automatically.

Related


Want this applied to your own LLM spend? FinOps LLM runs a free audit of your AI costs and shows where the savings are. Book free audit →

Back to research