# OpenAI GPT-5.6 Sol Cuts Factual Errors by 68 Percent in Tiered Release

> The GPT-5.6 family introduces Sol for complex reasoning with major factuality gains alongside Luna for efficiency that reaches 59.6 percent on ARC-AGI-2 at sharply reduced costs after July 2026 adjustments.

*Published 2026-08-07 · By Marcus Vance*

The GPT-5.6 family is OpenAI's set of frontier models released for general availability on July 9, 2026, featuring Sol, Terra, and Luna tiers.

OpenAI launched the GPT-5.6 family following a limited preview period that allowed initial testing before broader access. The models target different segments of the AI market with specialized optimizations for accuracy, speed, and expense.

## Background and Context of the Release

OpenAI has refined its model lineup across successive generations to meet growing demands for reliable outputs in professional settings. The GPT-5.5 Instant established a prior baseline that the new family surpasses in targeted evaluations.

Tiered releases allow organizations to match model capabilities to specific workloads without overpaying for unused performance. This approach reflects broader industry shifts toward flexible deployment options that scale with task complexity.

The July 9, 2026, general availability followed internal previews that gathered feedback on integration and reliability. Company documentation outlines the family as a direct response to user needs for both high-stakes accuracy and routine efficiency.

## Details of the GPT-5.6 Family Release

The family comprises three models released together on July 9, 2026. Sol serves as the flagship optimized for complex reasoning, coding, science, and cybersecurity according to OpenAI at https://openai.com/index/gpt-5-6/.

Terra provides a balanced option suited to everyday work tasks. Luna focuses on speed and cost efficiency as the most affordable tier in the lineup.

Price adjustments announced on July 30, 2026, further lowered costs for Luna and Terra. OpenAI stated at https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/ that Luna would cost 80 percent less while Terra would cost 20 percent less.

## Technical Specifics and Benchmark Performance

GPT-5.6 Sol demonstrated clear gains in factuality during testing. In a high-stakes evaluation covering finance, medicine, and law, the model produced 68 percent fewer responses with factual errors than GPT-5.5 Instant, as reported by OpenAI at https://x.com/OpenAI/status/2085434713821565297.

GPT-5.6 Luna reached 59.6 percent on ARC-AGI-2 Semi-Private at maximum reasoning effort under the updated pricing, according to ARC Prize at https://arcprize.org/results/openai-gpt-5-6-luna-2026-07-30. The score came at a cost of 0.18 dollars per task.

Comparison of GPT-5.6 Model Tiers Based on OpenAI and ARC Prize DataModel VariantPrimary FocusFactuality ImprovementARC-AGI-2 ScorePricing ChangeSolComplex reasoning, coding, science, cybersecurity68% fewer errors vs GPT-5.5 InstantNot reportedNo specific change notedTerraEveryday workNot specified in evaluationsNot reported20% lessLunaEfficiency and routine tasksNot specified in evaluations59.6%80% less

The tiered naming convention pairs generation numbers with capability descriptors. This structure clarifies expectations for users selecting models for different operational requirements.

- Assess Sol for applications requiring minimal factual deviations in regulated sectors.
- Deploy Luna for high-volume routine tasks to control expenses.
- Combine Luna with higher-tier models for hybrid workflows that optimize overall spend.
- Track subsequent ARC Prize updates for additional benchmark validations on the new pricing.

## Market and Stakeholder Implications

The pricing shifts alter the economics of AI deployment for enterprises and developers. Reduced costs for Luna expand access to capable models without proportional increases in budget.

Stakeholders in finance, medicine, and law gain from the documented error reductions in Sol. These improvements support compliance and decision-making processes that depend on accurate information.

The overall price-performance advance may accelerate adoption across industries that previously viewed advanced models as cost-prohibitive. OpenAI positions the family as scalable intelligence that matches varying ambition levels.

## Expert Reactions to the GPT-5.6 Models

> GPT‑5.6 Luna is an ideal pair programmer for larger models, handling much of the routine work while striking an ideal balance between cost and intelligence.Walden Yan, Co-Founder and Chief Product Officer, Cognition

The comment from Walden Yan highlights practical pairing strategies that leverage Luna for initial work before escalation to flagship models. Such pairings address common workflow bottlenecks in software development and research.

Reactions from the community emphasize the dual focus on accuracy and affordability. The combination addresses longstanding challenges in scaling AI use while maintaining output quality.

## What's Next for the Frontier Models

OpenAI is expected to iterate on the family based on deployment data and benchmark feedback. Continued refinements could further close gaps in specialized domains.

Users and organizations should monitor integration patterns and emerging use cases. The July 2026 pricing changes establish a new reference point for evaluating future releases in the sector.

The emphasis on verifiable performance metrics like those from ARC Prize provides a framework for ongoing comparisons. This data-driven approach supports informed decisions about model selection over time.

## Sources

1. [We’re launching the GPT‑5.6 family of models for general availability following our limited preview: our new flagship, Sol, alongside Terra, a balanced model for everyday work, and Luna, our most cost-efficient model.](https://openai.com/index/gpt-5-6/)
2. [Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, while GPT‑5.6 Terra, our balanced model for everyday work, will cost 20% less.](https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/)
3. [On July 30, 2026, OpenAI announced an 80% price reduction for GPT-5.6 Luna. At max reasoning effort under the new pricing, Luna scores 90.7% on ARC-AGI-1 Semi-Private at $0.07/task and 59.6% on ARC-AGI-2 Semi-Private at $0.18/task.](https://arcprize.org/results/openai-gpt-5-6-luna-2026-07-30)
4. [In our high-stakes factuality evaluation covering finance, medicine and law, the new GPT‑5.6 Sol produced 68% fewer responses with factual errors than GPT ‑5.5 Instant.](https://x.com/OpenAI/status/2085434713821565297)
5. [We find that GPT-5.6 Sol makes slightly fewer factual errors than GPT-5.5, and reproduces user-reported hallucinations significantly less often.](https://deploymentsafety.openai.com/gpt-5-6)

---
Source: https://aiintelreport.com/frontier-models/openai-gpt-5-6-sol-luna-release
Index: https://aiintelreport.com/llms.txt · Full text: https://aiintelreport.com/llms-full.txt
