ChatGPT
OpenAI's assistant and API, now on the GPT-5.6 Sol, Terra and Luna tiers, with a 1M token context window on all three and reasoning effort you set per request.
Overview
GPT-5.6 landed on 9 July 2026 in three tiers. Sol is the coding and reasoning flagship, Terra is the balanced default and Luna is the volume tier. All three take a 1 million token context window on the API and all three accept a reasoning effort setting on every request, which is a cleaner design than maintaining separate reasoning models. Sol adds an Ultra mode that runs four parallel subagents.
Pricing moved hard on 30 July. Luna fell 80% to $0.20 per million input and $1.20 output, Terra fell 20% to $2 and $12, and Sol held at $5 and $30. Batch takes another 50% off and cached prefix reads bill at 10% of rate, so a pipeline with a stable system prompt lands well under the sticker price. On data handling, the API does not train on customer content by default, standard retention is 30 days and Zero Data Retention was extended to frontier models for eligible accounts on 19 August 2026.
The obvious use is agentic work: long tool-calling chains, terminal tasks, browser automation, anything where the model has to stay coherent across dozens of steps. Sol's 88.8% on Terminal-Bench 2.1 shows up in practice as fewer dead-ended runs. Teams doing bulk document analysis on a budget should price Gemini 3.1 Pro at $2 input before committing, and anyone reaching for Luna on long documents should look at its 41.3% MRCR recall first. One practical note for teams inside mainland China: this is not officially served there, so the deployment story runs through overseas infrastructure.
Key Features
- ✓ GPT-5.6 shipped 9 July 2026 in three tiers: Sol for coding and reasoning, Terra as the balanced default, Luna for volume
- ✓ 1 million token context window on all three API tiers, with 128K maximum output
- ✓ Reasoning effort configurable per request from none through to max, plus a Sol Ultra mode that fans out four parallel subagents
- ✓ Prompt caching at a 90% discount on cached reads with a 30 minute minimum, Batch at 50% off, Flex at half rate and Priority at 2.5x
- ✓ Built in tools callable from the Responses API: web search, code execution and file search
- ✓ Zero Data Retention for eligible API customers, extended to frontier models on 19 August 2026
Where it holds
- • Best agentic numbers currently published: Sol at 92.2% on BrowseComp, 88.8% on Terminal-Bench 2.1 and 62.6% on OSWorld 2.0
- • Luna at $0.20 input after the 80% cut makes bulk classification and extraction work that used to be marginal obviously worth running
- • Per-request reasoning effort is the most practical cost lever any frontier API offers right now
- • API does not train on customer content by default, 30 day standard retention and ZDR available to enterprise accounts on request
Where it breaks
- • The 1M context is an API property, not a ChatGPT one. Reasoning context caps at 400K on the $200 Pro seat and 256K on Plus, so the headline number is not what you get in the app.
- • Sol at $5 input is 2.5x Gemini 3.1 Pro's rate for a lot of work where Gemini is close enough.
- • Ultra mode runs roughly 3x standard Sol pricing for a typical 3 to 4 point benchmark gain. Hard to justify outside a narrow band of problems.
- • Luna falls off a cliff on long context at 41.3% MRCR recall against Sol's 91.5%, so the cheap tier is not a drop-in for document work.
My Take
The 30 July price cut reset the maths. Luna dropped 80% to $0.20 per million input and Terra 20% to $2, which pushes a lot of extraction and classification work from marginal into obviously worth doing. Sol still holds the agentic crown at 92.2% on BrowseComp, and the per-request reasoning dial from none to max is the most useful cost lever in any frontier API. The catch is that the 1M context window is an API property, not a ChatGPT one: Pro reasoning caps at 400K, so the $200 seat does not buy the long context the launch material advertises.
Quick Info
- Pricing:
- freemium
- Openness:
- Proprietary
- Starting at:
- Free tier; Go $8/mo; Plus $20/mo; Pro at $100/mo (5x limits) and $200/mo (20x limits); Business $25/user/mo monthly or $20/user/mo annual with a two seat minimum, plus premium Business seats at $125/user/mo announced 10 August 2026. API per million tokens after the 30 July 2026 cut: Sol $5 in / $30 out, Terra $2 / $12, Luna $0.20 / $1.20. Batch API takes a flat 50% off, cached prefix reads bill at 10% of rate.
- Added:
- Aug 2026
- Updated:
- Aug 2026
Use Cases
Judge it on your own work
The notes above say where ChatGPT holds and where it breaks. The fastest check is your own workload.
Visit website ↗Alternatives to ChatGPT
Claude
freemiumAI assistant focused on safety, accuracy, and nuanced understanding for complex tasks and analysis.
Google Gemini
freemiumGoogle's assistant and API. Gemini 3.1 Pro carries a 1,048,576 token context on every tier and the lowest frontier input price at $2 per million.
DeepSeek
freemiumChinese AI lab producing open-source LLMs that rival proprietary models at a fraction of the cost, with strong reasoning and coding abilities.
Kimi / Moonshot AI (月之暗面)
freemiumMoonshot AI’s long-context AI assistant, pioneering ultra-long document processing with a 2 million token context window.