Observed Signal · Jul 30, 2026 · Technical Release · Source: OpenAI Blog · Impact: 4/5 · Sentiment: Positive

OpenAI Cuts GPT-5.6 Prices, Adds Fast Mode

Executive Signal Summary

OpenAI announced sizable price cuts and a new performance tier for its GPT-5.6 family roughly three weeks after the models’ release. Mid-tier Terra is down 20% to $2 per million input tokens and $12 per million output tokens; fastest Luna is down 80% to $0.20 per million input and $1.20 per million output. The company also introduced Fast mode (replacing Priority Processing) for Sol, offering up to 2.5× faster generation at 2× the price, while Sol’s base pricing remains unchanged. OpenAI said the changes reflect engineering gains that improve token-generation efficiency and serving costs and respond to enterprise cost sensitivity and competitive pressure. All three models (Sol, Terra, Luna) remain available through the OpenAI API, ChatGPT Work and Codex, with subscription-tier access rules for Terra and Luna.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

Lower prices and a faster API processing tier from a major LLM provider materially reduce cost and latency for integrating advanced models into marketing, analytics, and adtech workflows, enabling broader enterprise adoption.

SIGNAL RADAR

Track OpenAI Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • GPT-5.6 Terra price cut 20% to $2 per million input tokens and $12 per million output tokens.
  • GPT-5.6 Luna price cut 80% to $0.20 per million input tokens and $1.20 per million output tokens.
  • OpenAI offers three GPT-5.6 models — Sol (highest capability, pricing unchanged), Terra (mid-tier), and Luna (fastest).
  • Fast mode replaces Priority Processing for Sol, delivering up to 2.5× speed at 2× the price.
  • All models remain available via the OpenAI API, ChatGPT Work and Codex; Terra is accessible to Free/Go users and Terra/Luna are selectable by Plus, Pro, Business and Enterprise tiers; changes cite engineering gains and competitive/enterprise cost pressures.
Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: OpenAI Blog•Published: Jul 30, 2026
Original Coverage Title: “Advancing the price-performance frontier with GPT-5.6”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

Large Language Models (LLM) & AIJun 26, 2026

OpenAI previews GPT‑5.6 Sol, Terra and Luna

OpenAI previewed its GPT-5.6 family (including Sol, Terra and Luna) and began a limited release on 2026-06-26, restricting access to a small group of government-approved partners at the U.S. administration’s request. GPT-5.6 Sol is presented as the most capable variant and — according to reporting referenced in this piece — tests nearly as powerful yet cheaper to use than the so-called Mythos-class models. The U.S. government has treated both this release and Mythos-class models as high-risk and is coordinating controlled access while it develops a repeatable release process. The author characterizes this moment as “unprecedented,” warning that these models’ growing agentic capabilities increase risks (including cybersecurity concerns) and will accelerate disruptive change across industries.

Read assessment
Large Language Models (LLM) & AIJul 9, 2026

OpenAI launches GPT-5.6 family (Sol, Terra, Luna)

OpenAI announced the general-availability launch of the GPT-5.6 family—flagship Sol, balanced Terra, and cost-efficient Luna—available in ChatGPT (renamed ChatGPT Classic), ChatGPT Desktop (formerly Codex), and the OpenAI API. GPT-5.6 Sol claims state-of-the-art results across coding, knowledge work, cybersecurity, and science while improving efficiency versus prior frontier models. New capabilities include an Ultra multi-agent Sol tier that coordinates four agents in parallel, a Max reasoning mode, and Programmatic Tool Calling in the Responses API; OpenAI also introduced prompt-caching changes and ZDR-compatible in-memory tool execution. The release includes layered safeguards, extensive red-teaming, and Trusted Access for sensitive defensive uses; access to the most capable cybersecurity models will require hardware-based passkeys by 2026-09-01. OpenAI published per-1M-token pricing for Sol, Terra, and Luna and cited benchmark leads versus Anthropic's Claude Fable 5.

Read assessment
Large Language Models (LLM) & AIJul 11, 2026

OpenAI's GPT-5.6: Sol, Terra, Luna Tier Comparison

The author tested OpenAI's GPT-5.6 family (Sol, Terra, Luna) under real traffic to compare cost, latency, and capability trade-offs. Official list prices per million tokens are presented for each tier (Sol, Terra, Luna). Terra proved to deliver roughly 97% of Sol's benchmark performance at about half the price, making it a practical default for everyday tasks. Luna is positioned as the low-cost, low-latency tier for high-volume classification and streaming use cases, while Sol is recommended for high-stakes, deep-reasoning workloads and complex agent orchestration ("ultra mode"). The article also highlights predictable caching: prompt prefixes can remain cached for at least 30 minutes and cache reads bill at 10% of the input price, a factor that can invert cost comparisons between tiers. The author recommends routing rules: start on Terra, promote to Sol when needed, and use Luna for bulk/latency-sensitive lanes, testing all three behind a single API key (via byesu).

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.