OpenAI API pricing in 2026: GPT-5.6 family and beyond

Prices verified August 21, 2026 ยท USD per 1M tokens

OpenAI bills per million tokens with separate input, cached-input and output rates. The August 2026 lineup centers on the GPT-5.6 family โ€” Sol (flagship), Terra (balanced) and Luna (fast/cheap) โ€” with the GPT-5.4 family and GPT-4o still available for compatibility.

ModelInputCached inputOutputContext
GPT-5.6 Sol$5.00$0.50$30.001M
GPT-5.6 Terra$2.00$0.20$12.001M
GPT-5.6 Luna$0.20$0.02$1.201M
GPT-5.5 Pro$30.00โ€”$180.001M
GPT-5.4$2.50$0.25$15.001M
GPT-5.4 mini$0.75$0.075$4.50400K
GPT-5.4 nano$0.20$0.02$1.25400K
GPT-4o$2.50$1.25$10.00128K
GPT-4o mini$0.15$0.075$0.60128K

Cached input bills at ~10% of the standard input rate on GPT-5.x models. The Batch API takes 50% off both input and output on supported models.

Which GPT model is worth paying for?

Terra is the production workhorse โ€” GPT-5.6 quality at $2/$12 covers most chat, RAG and summarization loads. Luna at $0.20/$1.20 is one of the best price-performance deals on the market for classification, extraction and short generation. Sol earns its 2.5ร— premium on hard reasoning; GPT-5.5 Pro at $30/$180 is a specialist tool โ€” at 6ร— Sol's output price, reserve it for tasks where nothing else passes your evals.

The bill is mostly output tokens

Output costs 6ร— input across the GPT-5.6 family, so trimming max_tokens and verbose formatting usually saves more than switching models. Combine with automatic prompt caching (repeated prefixes bill at 10%) and the Batch API (โˆ’50%) before paying for a bigger tier โ€” worked examples in our caching guide.

โ†’ What would YOUR workload cost?
Enter your token mix once โ€” compare all 28 models side by side, with caching and batch discounts.
Advertisement

Frequently asked questions

How much does the GPT-5.6 API cost?
GPT-5.6 Sol costs $5 input / $30 output per million tokens, GPT-5.6 Terra $2/$12, and GPT-5.6 Luna $0.20/$1.20 (August 2026). Cached input is ~10% of the input rate.
What is the cheapest OpenAI model?
GPT-4o mini at $0.15/$0.60 per million tokens, followed by GPT-5.6 Luna at $0.20/$1.20 โ€” Luna is the newer and generally stronger choice.
How much does GPT-5.5 Pro cost?
$30 input / $180 output per million tokens โ€” roughly 6ร— GPT-5.6 Sol. A single long response can cost several dollars, so it is priced for high-stakes reasoning tasks.
Does OpenAI offer prompt caching?
Yes โ€” repeated prompt prefixes are cached automatically on GPT-5.x models and billed at about 10% of the normal input rate. No code changes required beyond keeping your prompt prefix stable.
How does OpenAI batch pricing work?
The Batch API processes requests asynchronously (up to 24h) at 50% off both input and output tokens โ€” ideal for evals, ETL and bulk generation.