# OpenAI Cuts GPT-5.6 Luna Prices by 80 Percent to $0.20 per Million Input Tokens

> Permanent reductions driven by inference efficiency position Luna for agentic workloads while GPT-5.6 Terra drops 20 percent and Google counters with Gemini 3.6 Flash at $1.50 input pricing.

*Published 2026-08-11 · By Marcus Vance*

GPT-5.6 Luna is OpenAI's fastest and most affordable frontier model after an 80 percent price reduction to $0.20 per million input tokens.

OpenAI announced an 80 percent reduction in the pricing of its GPT-5.6 Luna model as part of updates to the GPT-5.6 family. The change is effective July 30, 2026, and applies permanently to the API. This adjustment follows efficiency gains achieved in inference and serving processes.

## What prompted the pricing adjustments for GPT-5.6 models?

The primary driver behind the price cuts is the achievement of greater efficiency in how the models are run. OpenAI has indicated that these improvements allow for lower costs without the need for loss leading. The Luna model benefits most from this development as the fastest and most affordable option.

Prior to the update Luna pricing was higher and the reduction brings it to a level that supports broader adoption in cost sensitive scenarios. Terra also receives a reduction though smaller in scale. Sol sees no change but gains a fast mode feature for certain operations.

## What are the new prices for GPT-5.6 Luna, Terra, and Sol?

The new input token price for GPT-5.6 Luna is $0.20 per million. The output token price for the same model is $1.20 per million. These figures reflect the full 80 percent reduction applied to input tokens from earlier rates.

For GPT-5.6 Terra the input price is now $2 per million tokens. The output price is $12 per million tokens. This constitutes the 20 percent reduction from earlier pricing levels.

GPT-5.6 Sol pricing has not changed from its previous levels. Users of this model can now access a fast mode that delivers up to 2.5 times the speed in certain operations without any adjustment to base rates.

Comparison of API pricing for selected frontier modelsModelInput Price ($ per million tokens)Output Price ($ per million tokens)Price ChangeGPT-5.6 Luna0.201.20-80%GPT-5.6 Terra2.0012.00-20%GPT-5.6 SolUnchangedUnchanged0%Gemini 3.6 Flash1.507.50N/A

## How does Gemini 3.6 Flash factor into the pricing competition?

Google has introduced Gemini 3.6 Flash with input pricing set at $1.50 per million tokens. The output pricing for this model is $7.50 per million tokens. It is designed for agentic and coding workloads where token efficiency provides additional advantages.

This pricing allows Gemini 3.6 Flash to offer competitive costs for building and running agents. The model is presented as reducing the overall cost per agentic task through its design and efficiency gains.

> major price cuts today: *80% drop for GPT-5.6 Luna, now $0.20 per million input tokens and $1.20 per million output *20% drop for GPT-5.6 TerraSam Altman, CEO, OpenAI

## What market implications arise from these price reductions?

The lower pricing for Luna can lead to increased usage of agentic AI by lowering the barrier for frequent calls. Stakeholders such as developers and enterprises may find it more feasible to deploy larger scale agent systems.

The overall effect is an escalation in the price war for agent focused models. Both OpenAI and Google are adjusting their offerings to capture share in this growing segment of the market.

- Agent developers benefit from reduced costs per task when using the updated Luna pricing.
- Efficiency improvements provide a sustainable basis for ongoing price competitiveness.
- The addition of fast mode to Sol offers performance options without price increases.
- Competition with Gemini 3.6 Flash may lead to further optimizations by both companies.

## What comes next for frontier model pricing?

Further adjustments may occur as efficiency technologies advance in the industry. The focus on agentic workloads suggests continued emphasis on cost per task metrics across providers.

Companies will likely continue to highlight how their models achieve better price performance through technical means. This trend supports wider accessibility to advanced AI capabilities for a range of applications.

## Sources

1. [OpenAI announced an 80% reduction in GPT-5.6 Luna pricing to $0.20 per million input tokens and $1.20 per million output tokens effective July 30, 2026, along with a 20% reduction for GPT-5.6 Terra.](https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/)
2. [Sam Altman announced major price cuts for GPT-5.6 Luna and Terra.](https://x.com/sama/status/2082880720989532597)
3. [Gemini 3.6 Flash is priced at $1.50 per million input tokens and $7.50 per million output tokens for agentic workloads.](https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/)

---
Source: https://aiintelreport.com/frontier-models/openai-gpt-5-6-luna-price-cut
Index: https://aiintelreport.com/llms.txt · Full text: https://aiintelreport.com/llms-full.txt
