OpenAI now sells three GPT-6 models, Astra since September 3, 2026 and GPT-6 Sol and GPT-6 Luna since September 22, and none of them took over the ChatGPT conversation window. Chat still runs on GPT-5.6, public since July 9, 2026, and GPT-6 reaches it only as the Pro option on Pro, Business and Enterprise plans. The flagship Sol tier scores 88.8% on Terminal-Bench 2.1 (and 91.9% in its Sol Ultra high-effort mode), edging GPT-5.5’s own 88.0% on the same agentic-coding test, and it ships alongside two cheaper tiers, Terra and Luna, that undercut the previous generation on price.

In under a year, OpenAI has shipped seven distinct versions of GPT-5, each with its own identity and price point. If you have lost track of what separates GPT-5.0 from GPT-5.6, this guide covers every model in the family, release dates, context windows, benchmarks, pricing, and the key differences that actually matter. GPT-5.6 first shipped as a June 26 government-gated preview to roughly 20 organizations, then went fully public on July 9. It now powers the reasoning options on ChatGPT’s paid plans and is available in full through the API. We also cover where the three GPT-6 models land, and which of them your plan can actually open. If you are newer to ChatGPT, our ChatGPT for Beginners guide is a good place to start before you compare model differences.

The Key Takeaways

  • GPT-5.6 (public July 9, 2026) is what the ChatGPT conversation window still serves on every plan, Sol on paid plans and Luna on Free and Go. It ships in three tiers, Sol (flagship), Terra (balanced), and Luna (cheapest), with Sol scoring 88.8% on Terminal-Bench 2.1 and a 1.05M-token API context window.
  • GPT-5.5 (April 23, 2026) is the previous flagship and still an excellent all-rounder. It scores 93.6% on GPQA Diamond and 78.7% on OSWorld-Verified, both ahead of GPT-5.4, and ships in two variants, GPT-5.5 standard and GPT-5.5 Pro.
  • GPT-5.4 (March 5, 2026) sits two generations back but still matters. It introduced a 1M-token context window (API only) and native computer use, scoring 75.0% on OSWorld-Verified and 57.7% on SWE-Bench Pro.
  • GPT-5.3 Instant (March 3, 2026) fixed the over-cautious tone and delivered 26.8% fewer hallucinations than GPT-5.2 with web search on, but it is no longer listed on OpenAI’s API pricing page. Only gpt-5.3-codex survives from that generation.
  • GPT-5.2 was the first model to score 90%+ on ARC-AGI-1 (Pro) and hit a perfect 100% on AIME 2025 math.
  • GPT-5.1 introduced adaptive reasoning, 2 to 3x faster on simple tasks, while keeping the same $1.25 / $10 price as GPT-5.0.
  • Pricing splits by tier in GPT-5.6. Sol was cut to $4 / $20 per 1M tokens on August 21, 2026, below GPT-5.5’s $5 / $30, Terra drops to $2 / $12, and Luna is just $0.20 / $1.20 after OpenAI cut both tiers on July 30, 2026, Luna by 80% and Terra by 20%. OpenAI now calls Sol’s rate promotional and guarantees it only through November 21, 2026.
  • 15 legacy model snapshots lost API access on July 23, 2026, including five Codex snapshots and both chat-latest aliases. It is an API-only change; the ChatGPT model picker is untouched.
  • GPT-6 Astra shipped on September 3, 2026. It costs $10 / $50 per 1M tokens, roughly 2.5x GPT-5.6 Sol, and it reached ChatGPT the next day as GPT-6 Pro on Pro, Business and Enterprise. Plus plans get Astra in ChatGPT Work and Codex instead, and Free and Go get none.
  • GPT-6 Sol and GPT-6 Luna followed on September 22, 2026 at $2 / $10 and $0.10 / $0.50, half the matching GPT-5.6 rates. They run in ChatGPT Work, Codex and the API, and OpenAI’s release notes say plainly that they are separate from the models available in Chat.

All ChatGPT 5 Models at a Glance

Do editor

Todos os modelos de IA numa só aplicação

Fello AI reúne GPT-6, Claude 5, Gemini 3.8, Grok 4.7 e mais numa só aplicação nativa para Mac e iPhone.

Descarregue já!
Model Released Context Window API Pricing (Input / Output) Best For
GPT-5.0 Aug 7, 2025 400K / 128K out $1.25 / $10 per 1M General use, launch baseline
GPT-5.1 Nov 13, 2025 400K (272K in) $1.25 / $10 per 1M Adaptive speed, conversational tasks
GPT-5.2 Dec 11, 2025 400K (272K in) $1.75 / $14 per 1M Deep reasoning, coding, research
GPT-5.3 Instant Mar 3, 2026 400K ~$0.30 / ~$1.20 per 1M Everyday writing, cost-sensitive use
GPT-5.4 Mar 5, 2026 1M (API only) $2.50 / $15 per 1M Agentic work, computer use
GPT-5.5 Apr 23, 2026 1M (API only) $5 / $30 per 1M (Pro: $30 / $180) Agentic coding, knowledge work
GPT-5.6 Jul 9, 2026 1.05M (128K out) Sol $4/$20, Terra $2/$12, Luna $0.20/$1.20 Agentic coding, knowledge work: what ChatGPT Chat still serves

One caveat on that table. It is a history of the family, not a current order form. OpenAI’s API pricing page still carries most of these models, from gpt-5 and its mini, nano and pro variants up through the GPT-5.4 family, GPT-5.5, GPT-5.5 Pro, the three GPT-5.6 tiers and the three GPT-6 models, with gpt-5.3-codex listed separately under Codex. The standard GPT-5.3 is the one that has gone. What remains is on the clock: the original gpt-5-2025-08-07, along with its mini, nano and pro variants, shuts down on December 11, 2026. The prices below are kept because they explain how the family evolved, not because you can still buy them all.

The GPT-6 family now sits above every model in that table, and which part of it you can reach depends on the surface, not the price. GPT-6 Astra launched on September 3, 2026 at $10 / $50 per 1M tokens, with the same 1.05M-token context and 128K output ceiling, and GPT-6 Sol ($2 / $10) and GPT-6 Luna ($0.10 / $0.50) followed on September 22 at half the matching GPT-5.6 rates. Inside ChatGPT, Astra appears as the Pro option on Pro, Business and Enterprise, while Sol and Luna stay in Work, Codex and the API. Our GPT-6 Astra launch coverage has the full pricing and benchmark breakdown.

For how these API prices map to ChatGPT subscription tiers (Free, Go, Plus, Pro, Business, Enterprise), see our ChatGPT pricing guide.

Download Fello AI to use all the best AI models, work with files and generate images, available on iPhone, iPad and Mac

GPT-5.0: The Foundation

Released: August 7, 2025

The original GPT-5 arrived as a meaningful jump over GPT-4o, not just in raw benchmarks, but in architecture. According to OpenAI’s launch announcement, it was built as a unified system with a fast base model for everyday queries and a deeper reasoning layer (GPT-5 Thinking) that activates automatically when the query demands it. A real-time router decides which to use based on complexity, tool needs, and context, so you do not have to manage it manually.

GPT-5.0 Benchmarks

Benchmark GPT-5.0 score
AIME 2025 (math) 94.6% without tools; 100% with Python tools (Pro)
GPQA Diamond (PhD-level science) 89.4%
SWE-bench Verified (coding) 74.9%
Aider Polyglot (real-world coding) 88%
Humanity’s Last Exam 42%
Hallucinations Under 1% on open-source prompts; 1.6% on hard medical cases

GPT-5.0 shipped with a 400K input / 128K output context window and was available to all ChatGPT users, with Pro subscribers getting extended reasoning access.

API pricing: $1.25 per 1M input tokens / $10 per 1M output tokens

What Changed vs. GPT-4o

GPT-5 was roughly 45% less likely to hallucinate than GPT-4o with web search enabled. The unified routing eliminated the need to manually switch between chat and reasoning modes, a friction point many users had with the o-series models.

GPT-5.1: Faster Without Being Dumber

Released: November 13, 2025

Per OpenAI’s GPT-5.1 announcement, GPT-5.1 was not a capability leap, it was an efficiency upgrade. The headline feature was adaptive reasoning, where the model dynamically allocates compute based on query complexity. Ask it something simple and it answers 2 to 3x faster than the standard model. Ask it something complex and it switches to full reasoning mode. That reasoning mode is also why heavier models can feel laggy; our guide to why ChatGPT is so slow explains when the wait is the model and when it is something you can fix.

OpenAI also tuned the tone to be warmer and more conversational, dropping some of the rigid formality that made GPT-5.0 occasionally feel stiff. It used 30% fewer thinking tokens than its Codex variant while maintaining near-identical benchmark scores on most tasks.

GPT-5.1 Benchmarks

Benchmark or spec GPT-5.1
AIME 2025 94% (marginally lower than GPT-5.0’s 94.6%)
GPQA Diamond 87%
MMMU (multimodal) 85.4%
SWE-Bench Verified 76.3% (Codex-Max variant: 77.9%)
Context window 400K tokens (272K input / 128K output)
API pricing $1.25 per 1M input / $10 per 1M output, same as GPT-5.0

GPT-5.1 Variants

OpenAI shipped a full family with GPT-5.1, Instant (fast), Thinking (deep reasoning), Auto (routing), and Pro (research-grade), plus three Codex variants, standard Codex, Codex-Mini for lightweight tasks, and Codex-Max for agentic coding tasks lasting 24+ hours. All three of those Codex snapshots lose API access on July 23, 2026, with gpt-5.5 as the substitute for Codex and Codex-Max and gpt-5.4-mini for Codex-Mini.

What GPT-5.1 Got Right

The reasoning_effort parameter (none / low / medium / high) gave developers fine-grained control over how much compute to spend per request, a practical tool for balancing cost and quality in production.

GPT-5.2: The Reasoning Milestone

Released: December 11, 2025

GPT-5.2 was the model that set new goalposts. It was the first AI to score 90%+ on ARC-AGI-1, a benchmark designed specifically to resist pattern-matching and test genuine reasoning. It also hit a perfect 100% on AIME 2025 math problems. These were not marginal gains; they crossed thresholds that had defined the frontier for years.

OpenAI introduced a three-tier architecture with GPT-5.2. GPT-5.2 Instant was optimized for throughput, which made it the pick for customer support, content generation, and translation. GPT-5.2 Thinking added configurable reasoning depth through Light, Medium, Heavy, and xhigh settings, letting you trade latency for accuracy on a per-request basis. GPT-5.2 Pro sat at the top with maximum compute and up to 30 minutes of sustained processing for the most demanding tasks.

GPT-5.2 Benchmarks

Benchmark Score
AIME 2025 100%
GPQA Diamond (Pro) 93.2%
GPQA Diamond (Thinking) 92.4%
SWE-Bench Verified 80.0%
SWE-Bench Pro 55.6%
ARC-AGI-1 (Pro) 90%+
ARC-AGI-2 (Pro) 54.2%
FrontierMath 40.3%

GPT-5.2 delivered 38% fewer errors than GPT-5.1, the largest reliability jump in the GPT-5 family so far.

The context window stayed at 400K tokens (272K input / 128K output), but API pricing rose to $1.75 per 1M input / $14 per 1M output, a 40% increase over GPT-5.1.

An agentic coding variant, GPT-5.2-Codex, followed on January 14, 2026, purpose-built for planning and executing multi-step engineering tasks autonomously. It retires from the API on July 23, 2026 alongside the GPT-5.0 and 5.1 Codex snapshots.

Note: GPT-5.2 Thinking was retired on June 3, 2026. If you relied on it for analytical work, GPT-5.4 Thinking, GPT-5.5, or GPT-5.6 Sol is the intended upgrade path.

GPT-5.3: The Everyday Upgrade

Released: March 3, 2026

GPT-5.3 Instant fixed something earlier models quietly got wrong, the tone. Previous GPT-5 versions had a tendency toward excessive caveats, unnecessary hedging, and what users called “cringe,” AI overcaution that made every answer feel like a legal disclaimer. GPT-5.3 dialed that back, producing more direct and natural responses.

It also delivered a meaningful accuracy improvement. With web search enabled, GPT-5.3 produces 26.8% fewer hallucinations than GPT-5.2 Instant. Without search, the improvement is still 19.7%. User-flagged errors dropped by 22.5%.

GPT-5.3 Specs

Spec GPT-5.3 Instant
Context window 400K tokens
HealthBench 54.1% (slight dip from GPT-5.2’s 55.4%)
Hallucinations 26.8% fewer, with web search, versus GPT-5.2
API pricing ~$0.30 per 1M input / ~$1.20 per 1M output

That pricing was a dramatic drop from GPT-5.2’s $1.75 / $14, and GPT-5.3 Instant was positioned as a high-quality, low-cost everyday model rather than a reasoning powerhouse. It no longer appears on OpenAI’s API pricing page, so treat those figures as history. If you want that role today, GPT-5.6 Luna at $0.20 / $1.20 is the direct successor, and since OpenAI cut it 80% on July 30, 2026 it is now cheaper than GPT-5.3 Instant ever was.

GPT-5.3 Trade-offs

The anti-cringe update came with a trade-off: safety compliance on some categories declined. Graphic violence content filtering dropped from 85.2% to 78.1% compared to GPT-5.2, a decision OpenAI appears to have made deliberately, moving some safety controls to the product layer rather than the model level.

The GPT-5.3-Codex variant (February 5, 2026) is worth noting separately. It carries a 1 million-token context window, the same as GPT-5.4, runs 25% faster than GPT-5.2-Codex, and is the best option for large-scale agentic coding at a lower cost than 5.4 or 5.5. If you run any Codex variant locally, check our explainer on Codex’s storage bug first, because a runaway logger was burning through SSD endurance until OpenAI patched most of it in June 2026.

GPT-5.4: The Professional Frontier

Released: March 5, 2026

GPT-5.4 held the “most capable model OpenAI has ever shipped” title for six weeks until GPT-5.5 arrived in late April 2026. According to OpenAI’s official GPT-5.4 announcement, the two genuinely new capabilities that set it apart from every prior model were native computer use and a 1 million-token context window (via the API).

Computer use means GPT-5.4 can interact directly with software interfaces, navigating desktops, clicking UI elements, running commands, verifying output, and looping back to fix errors in a build-run-verify-fix cycle. On OSWorld-Verified, the benchmark for desktop computer use, it scores 75.0%, exceeding the measured human baseline of 72.4%.

GPT-5.4 Benchmarks

Benchmark GPT-5.4 GPT-5.2
OSWorld-Verified (computer use) 75.0% 47.3%
GDPval (knowledge work, 44 professions) 83.0% 70.9%
SWE-Bench Pro (coding) 57.7% 55.6%
GPQA (science) 92.8% 93.2%
ARC-AGI v2 73.3% 54.2%
FrontierMath 47.6% 40.3%
MMMU-Pro (multimodal) 81.2%
BrowseComp (web research) 82.7%
Spreadsheet tasks 87.3%
Hallucinations vs. GPT-5.2 33% fewer (individual claim errors)

GPT-5.4 Key Features

1 million-token context (API only). The full 1M window is available via the API. ChatGPT Plus, Team, and Pro subscribers in the chat interface do not get the expanded context. Worth noting, OpenAI charges double the standard rate once input exceeds 272K tokens in a single request.

Tool search. Instead of loading all tool definitions upfront, GPT-5.4 can retrieve tool definitions on demand. This cuts token usage by 47% in tool-heavy workflows, a practical cost saving for anyone building agents.

Full-resolution vision. GPT-5.4 processes images up to 10.24 million pixels, making it viable for detailed medical imaging, architectural plans, and high-res document analysis.

Compaction training. The model is trained to compress long agent trajectories while preserving key context, useful for multi-day autonomous workflows.

GPT-5.4 Pricing

Tier Input per 1M Output per 1M
Standard $2.50 $15.00
Batch / async $1.25 $7.50
Priority $5.00 $30.00
GPT-5.4 Pro $30.00 $180.00

GPT-5.4 Pro is priced at a significant premium. For context, Anthropic’s Claude Opus 4.8 costs $5 per 1M input / $25 per 1M output, so GPT-5.4 Pro is six times the input price.

GPT-5.4 ships in three variants: Standard (everyday professional use), Thinking (deep multi-step reasoning, available to Plus/Team/Pro subscribers), and Pro (maximum performance).

GPT-5.5: The Previous Flagship

Released: April 23, 2026

GPT-5.5 held the flagship slot from late April until GPT-5.6 arrived in July, and it has not gone anywhere. Its fast variant, GPT-5.5 Instant, is still the default model for quick everyday responses on every ChatGPT plan, including Free. It arrived 49 days after GPT-5.4, making it the fastest flagship turnover in the GPT-5 family. Read OpenAI’s official GPT-5.5 announcement for the full technical brief, and our own GPT-5.5 launch coverage for the complete benchmark tables and real-world testing.

Two variants ship. GPT-5.5 standard is available to all paid ChatGPT tiers and Codex, and GPT-5.5 Pro is restricted to Pro, Business, and Enterprise tiers. OpenAI positions GPT-5.5 as a model “built specifically for real work and for powering agents,” with the largest efficiency gains going to agentic coding and long-horizon knowledge work.

GPT-5.5 Benchmarks

Benchmark GPT-5.5 GPT-5.5 Pro
Terminal-Bench 2.0 82.7%
GDPval (knowledge work) 84.9%
OSWorld-Verified (computer use) 78.7%
SWE-Bench Pro (coding) 58.6%
GPQA Diamond (science) 93.6%
BrowseComp (web research) 84.4% 90.1%
FrontierMath Tier 4 35.4% 39.6%
Humanity’s Last Exam (no tools) 41.4% 43.1%
CyberGym 81.8%

Versus GPT-5.4, the standout gains are OSWorld-Verified (+3.7 points), GPQA Diamond (+0.8), SWE-Bench Pro (+0.9), and Terminal-Bench 2.0 (+7.6). The numbers understate how the model feels to use, though. OpenAI reports that at NVIDIA, debugging cycles “dropped from days to hours” using GPT-5.5’s agentic workflow, which is the kind of real-world win that does not show up in a benchmark table.

GPT-5.5 Key Features

Agentic coding by default. GPT-5.5 ships tuned for multi-step engineering workflows. The same Codex tool that used to need supervision can now chain apply_patch, terminal commands, and verification steps over longer horizons without losing the thread.

Computer use, now production-grade. 78.7% on OSWorld-Verified puts GPT-5.5 firmly above the 72.4% human baseline. Desktop automation workflows that were flaky on 5.4 start to look viable on 5.5.

Lower hallucination rate. OpenAI reports meaningfully fewer hallucinations versus GPT-5.4 across standard knowledge-work tasks, though the company has not released a single headline figure.

API access followed the app. At launch, GPT-5.5 went live inside ChatGPT and Codex first, with API access arriving shortly afterwards under different safeguards than the ChatGPT rollout. Both surfaces have been generally available for months.

GPT-5.5 Pricing

Tier Input per 1M Output per 1M
GPT-5.5 Standard $5.00 $30.00
GPT-5.5 Pro $30.00 $180.00

Standard is exactly 2x the price of GPT-5.4 standard. GPT-5.5 Pro matches GPT-5.4 Pro’s $30 / $180, so if you already pay for Pro, the jump to 5.5 Pro is effectively free at the API level.

GPT-5.6: What ChatGPT Chat Actually Serves (Sol, Terra, Luna)

Released: July 9, 2026 (June 26, 2026 government-gated preview)

GPT-5.6 is the model the ChatGPT conversation window actually serves, on free and paid plans alike, and it is the release that reorganized the lineup into three named tiers instead of a “standard plus Pro” split. It shipped first as a June 26 preview to roughly 20 approved organizations under the new federal review regime for frontier models, then went fully public on July 9, and it is still what Chat serves across every tier, alongside Codex and the API. All three tiers share a 1.05M-token context window and 128K max output.

The three tiers map cleanly to budget and capability. Sol is the flagship, tuned for agentic coding and long-horizon work, with a high-effort Sol Ultra mode on top. Terra is the balanced middle option, competitive with GPT-5.5 at 60% less. Luna is the fastest and cheapest, aimed at high-volume everyday use. On Terminal-Bench 2.1, the agentic-coding benchmark, Sol scores 88.8% and Sol Ultra 91.9%, edging GPT-5.5’s 88.0% on the same test.

Tier Input per 1M Output per 1M Role
Sol (+ Sol Ultra) $4.00 $20.00 Flagship, agentic coding, hardest reasoning
Terra $2.00 $12.00 Balanced, ~GPT-5.5 quality at 60% less
Luna $0.20 $1.20 Fastest and cheapest, high-volume tasks

How GPT-5.6 Actually Appears in ChatGPT

The API sells Sol, Terra, and Luna by name. Inside ChatGPT the same models sit behind thinking levels, which trips people up. GPT-5.6 Sol powers Instant on eligible paid plans as well as the Medium, High, and Extra High levels, while GPT-5.6 Luna is the default on Free and Go. GPT-5.6 Sol Pro powers the Pro option, which on Pro, Business and Enterprise also offers GPT-6 Pro. Extra High is the ChatGPT-side name for the same maximum-effort mode that OpenAI benchmarks as Sol Ultra.

ChatGPT plan Medium and High Extra High Pro
Free and Go Not included Not included Not included
Plus Included Not included Not included
Pro Included Included Included
Business Included Included Included
Enterprise Included Included Included

Two consequences are worth knowing before you pay for anything. Terra is not selectable in a normal ChatGPT conversation at all, and Luna only shows up there as the Free and Go default; otherwise both live in Work, Codex, and the API, where Work serves all three tiers to Plus and above and Codex adds Terra for Free and Go. And since September 14, 2026, ChatGPT no longer switches from Instant to a thinking level on its own for Plus and Pro, so asking it to think harder does nothing. Pick the level yourself, and if you exhaust a GPT-5.6 thinking allowance, OpenAI says ChatGPT may continue on another available model rather than cutting you off, so check which model answered before you trust a long analytical reply.

Those prices are the short-context rates, and OpenAI now labels Sol’s rate promotional, guaranteed only through November 21, 2026. Push a single request past OpenAI’s long-context threshold and OpenAI bills the whole request at 2x input and 1.5x output, taking Sol to $8 / $30, Terra to $4 / $18, and Luna to $0.40 / $1.80. Budget for the tier you will actually hit.

One caveat worth flagging: OpenAI has not published GPT-5.6 numbers for the standard benchmarks it used with earlier models (no SWE-bench Verified, GPQA, or OSWorld figures for Sol yet), and independent evaluator METR noted the family is unusually good at benchmark-style tasks, so treat the Terminal-Bench lead as one data point rather than a full picture. For the complete breakdown of the tiers, benchmarks, and rollout, see our GPT-5.6 Sol, Terra, and Luna explainer. Once you pick a tier, our GPT-5.6 prompting guide walks through seven tips for getting sharper answers.

What OpenAI Retired on July 23, 2026

Shipping seven versions in eleven months leaves debris, and OpenAI cleared it. On July 23, 2026, 15 model snapshots lost API access in a single wave, announced back on April 22 under the heading “Legacy GPT model snapshots.” The reasoning OpenAI gave was straightforward, to improve reliability and make it easier to pick the right model.

Read the scope carefully. This was an API-only shutdown, and it removed nothing from the ChatGPT model picker. If you use ChatGPT through the app or the web, nothing changed for you on July 23. If you call the API directly, or you run a tool that pins a specific model snapshot, that deadline has already passed.

The Codex Snapshots Take the Biggest Hit

Five pre-5.3 Codex snapshots retire together, which makes this the largest single cut to the Codex lineup so far. The two chat-latest aliases go the same day, and so do both deep-research snapshots and the original computer-use preview. The table below covers the entries most likely to appear in a GPT-5-era codebase.

Retiring July 23, 2026 OpenAI’s substitute
gpt-5-codex gpt-5.5
gpt-5.1-codex gpt-5.5
gpt-5.1-codex-max gpt-5.5
gpt-5.1-codex-mini gpt-5.4-mini
gpt-5.2-codex gpt-5.5
gpt-5-chat-latest gpt-5.5
gpt-5.1-chat-latest gpt-5.5
o3-deep-research-2025-06-26 gpt-5.5-pro
o4-mini-deep-research-2025-06-26 gpt-5.5-pro
computer-use-preview-2025-03-11 gpt-5.4-mini

The remaining five are GPT-4o-era leftovers rather than GPT-5 models, covering the search-preview pair, a text-to-speech snapshot, and dated audio and realtime mini snapshots. Note what is not on the list. GPT-5.3-Codex survives, which matters because it carries a 1M-token context window and remains the cheaper option for large-codebase agentic work. GPT-5.4, GPT-5.5, and GPT-5.5 Pro all stay on sale too.

The Bigger Wave Lands in October

July 23 is the smaller of the two deadlines. On October 23, 2026, OpenAI shuts down a much more recognisable set, including gpt-3.5-turbo-0125, gpt-4-0613, gpt-4-turbo, gpt-4o-2024-05-13, o1, o1-pro, o3-mini, o4-mini, and the original gpt-image-1, along with fine-tuned models built on GPT-3.5 Turbo and GPT-4. Anyone still running a fine-tune on those base models has three months to retrain, which is the part most teams underestimate. A third wave follows on December 11, 2026, taking the original GPT-5 snapshots with it, gpt-5-2025-08-07 plus its mini, nano, and pro variants, along with o3 and o3-pro. Separately, the Videos API and every Sora 2 snapshot sunset on September 24, 2026. You can track every date on OpenAI’s deprecations page.

Which ChatGPT Model Should You Use?

The right pick depends entirely on the job, from cheap everyday chat to top-tier reasoning and coding, and the table below maps common tasks to the best-fit model. If you are weighing Google’s lineup instead, our Gemini model comparison runs the same version-by-version breakdown for Flash, Pro, and Flash-Lite.

Task Best Model Why
Everyday writing, email, summaries GPT-6 Luna, or GPT-5.6 Luna in Chat Cheapest OpenAI tier at $0.10 / $0.50, half the GPT-5.6 Luna rate for the same index score
High-volume API (cost matters) GPT-6 Luna At $0.10 / $0.50 it undercuts GPT-5.6 Luna and the legacy GPT-5.4-nano on both sides, with current-generation quality on top
Deep analytical reasoning GPT-6 Sol, or GPT-5.6 Sol in Chat GPT-6 Sol costs half of GPT-5.6 Sol and edges it on Artificial Analysis’s Intelligence Index v4.3.2, 47.5 against 47.0
Complex coding, bug fixing GPT-6 Sol in Codex Same generation as Astra at $2 / $10, and GPT-5.6 Sol’s 88.8% on Terminal-Bench 2.1 is still the GPT-5 family high
Agentic, multi-step automation GPT-6 Astra 74.1% on DeepSWE v1.1 against GPT-5.6 Sol’s 70.8%, in roughly half the time per task
Computer use / desktop automation GPT-6 Astra 72.6% on OSWorld V2-Offline, while GPT-5.5’s 78.7% is on the older OSWorld-Verified and is not comparable
Long document analysis (1M+ tokens) GPT-6 or GPT-5.6 (API) Both generations carry a 1.05M-token context window
Enterprise / maximum accuracy GPT-6 Pro (Astra) The highest option in the ChatGPT picker, on Pro, Business and Enterprise
Balanced quality at lower cost GPT-5.6 Terra ~GPT-5.5 quality at 60% less ($2 / $12), and OpenAI has not announced a GPT-6 Terra

For a broader comparison including Claude, Gemini, and other leading models, see our Best AI Models hub. For a head-to-head on OpenAI versus Anthropic specifically, check our Claude vs ChatGPT comparison. For Google’s side, see our Gemini 3.5 review and Pro release tracker. And if you would rather talk than type, our ChatGPT live voice mode guide explains OpenAI’s new real-time GPT-Live.

GPT-6 Astra: What Actually Shipped

The wait ended on September 3, 2026, when OpenAI launched GPT-6 Astra and settled the argument over whether Astra would arrive as GPT-6 or as another point release inside the GPT-5 line. It is the first model OpenAI has ever designated Critical for cybersecurity capability under its Preparedness Framework, which is why the rollout started with vetted defenders rather than with the model picker. That gate lasted exactly one day, because Astra reached ChatGPT as GPT-6 Pro on September 4 for Pro, Business and Enterprise plans. Company president Greg Brockman told a press briefing, “Welcome to the AGI era.” Our GPT-6 Astra launch coverage works through the pricing, the benchmarks, and the footnotes those numbers arrive with. Sam Altman had said GPT-6 would come sooner than the 2.5-year gap between GPT-4 and GPT-5, and at 13 months after GPT-5.0 that held.

What GPT-6 Astra Brings

Notes that survive the context window. The long-rumoured persistent memory arrived in a narrower form than the leaks suggested. Astra can keep notes across context windows and search back through earlier messages and tool output, instead of relying on compaction that summarises old work and can discard the one detail an agent needs later. It is experimental behind a config.toml setting at launch, and OpenAI says it becomes the default for Astra in the coming weeks.

Long-horizon computer use. This is where the jump is easiest to see. On OSWorld V2-Offline, OpenAI reports Astra at 72.6% in about 40 minutes per task against GPT-5.6 Sol’s 65.7% in about 75 minutes, and on DeepSWE v1.1 agentic coding, 74.1% against Sol’s 70.8%. Halving the time per task while raising the score is the practically useful result. Astra can also ask you a question without stopping the work that does not depend on your answer.

Scope discipline, at a cost. In OpenAI’s internal tests on difficult or impossible tasks run without production safeguards, GPT-5.6 Sol went beyond its authorised scope 48.2% of the time. Astra did so 0% of the time. The same materials disclose a regression: in evaluations designed to elicit monitoring evasion, Astra’s written reasoning was harder to follow than Sol’s. A model that never oversteps but whose reasoning is harder to read is a trade, not a clean win.

What did not arrive. The speculation around persistent weight updates, where a model learns across sessions rather than only within them, is not part of what OpenAI shipped. Neither is a GDPval figure, the benchmark OpenAI built to measure performance on economically valuable work, which is a conspicuous absence at a launch framed around AGI. And the headline 98.6% on ARC-AGI-3 is not a solve rate. It measures Relative Human Action Efficiency, on a harness OpenAI has itself shown can triple the number.

For how each GPT-6 rumour held up against what actually shipped, see our ChatGPT 6 leak scorecard.

Astra was not the whole family. On September 22, 2026 OpenAI added GPT-6 Sol at $2 / $10 and GPT-6 Luna at $0.10 / $0.50, halving the matching GPT-5.6 rates, and put both into ChatGPT Work and Codex for Plus and above, with Luna in the desktop app for Free and Go. OpenAI’s own release notes say the two are separate from the models available in Chat. Our guide to GPT-6 Sol and Luna covers the benchmarks and the surfaces in detail.

Benchmark Progression Through the GPT-5 Family

Benchmark GPT-5.0 GPT-5.1 GPT-5.2 GPT-5.3 GPT-5.4 GPT-5.5
AIME 2025 (math) 94.6% 94% 100% 100%
GPQA Diamond 89.4% ~87% 93.2% 92.8% 93.6%
SWE-Bench Verified 74.9% 76.3% 80.0% ~52.8%*
SWE-Bench Pro 55.6% 57.7% 58.6%
ARC-AGI-1 86.2% 93.7%
ARC-AGI-2 17.6% 54.2% 73.3%
FrontierMath (T1–3) 26.6% 40.3% 47.6%
FrontierMath (T4) 27.1% 35.4% (Pro 39.6%)
MMMU 84.2% 85.4% 84.2%
MMMU-Pro 86.5% 81.2%
HealthBench Hard 46.2% 55.4% 54.1% 62.6%
GDPval (knowledge work) 70.9% 83.0% 84.9%
OSWorld (computer use) 47.3% 75.0% 78.7%
BrowseComp 65.8% 82.7% 84.4% (Pro 90.1%)
Terminal-Bench 2.0 58.1% 64.9% 75.1% 82.7%
Spreadsheet modeling 68.4% 87.3%
HLE (Humanity’s Last Exam) 42% 52.1% 41.4% (no tools)
LiveCodeBench 72.5%
Aider Polyglot 88%
CyberGym 81.8%
Hallucination reduction -45% vs GPT-4o -38% vs 5.1 -26.8% vs 5.2 -33% vs 5.2 lower than 5.4 (no single figure yet)

GPT-5.4 SWE-Bench Verified score is from a non-thinking variant; OpenAI has not published an official figure for this benchmark. Humanity’s Last Exam numbers are not directly comparable, GPT-5.5’s 41.4% is the no-tools figure; GPT-5.4’s 52.1% is with tools.

The table stops at GPT-5.5 on purpose. The only headline benchmark OpenAI published for GPT-5.6 at launch is Terminal-Bench 2.1, where Sol hits 88.8% and Sol Ultra 91.9%. Two further Sol figures arrived sideways in the GPT-6 Astra launch materials on September 3, 2026, 70.8% on DeepSWE v1.1 and 65.7% on OSWorld V2-Offline, but both sit on benchmark releases the older rows never used, so adding them would mean comparing different tests. Comparable GPQA, SWE-Bench, and FrontierMath figures for the 5.6 family are still unpublished. Our guide to AI benchmarks explains what Terminal-Bench actually measures.

Want Every GPT-5 Variant Plus Claude, Gemini, and Grok in One Mac App?

If you are paying for ChatGPT, Claude Pro, and Google AI Pro just to compare model outputs, there is a simpler setup. Fello AI is a native Mac, iPhone and iPad app that routes your prompts to ChatGPT, Claude, Gemini, Grok, DeepSeek, Perplexity, Kimi, GLM and Qwen through one interface, for $9.99/month. One price, every top model, no tab switching.

That matters in a world where OpenAI ships a new GPT-5 variant every six weeks, and where each one competes head-to-head with releases from Anthropic, Google, and xAI. If GPT-5.6 wins on agentic coding today but Claude or Gemini edges ahead on a specific task next month, you can switch models inside the same app without changing your workflow.

You can download Fello AI on the App Store and run every one of these models side by side on Mac, iPhone, and iPad, for $9.99/month, with a free tier to try first and a 4.7-star rating across 27,000+ reviews. It works as an alternative to stacking three separate subscriptions, or as a complement to the one you already pay for, and it handles images, documents, and spreadsheets rather than chat alone.

For context on persistent memory, see Anthropic’s new agent memory system.

For practical tips on how to use ChatGPT effectively, see our full guide.

Conclusion

The GPT-5 family has moved fast. In eleven months, OpenAI has gone from a capable general-purpose model to one that operates software like a human, holds an entire codebase in context, and outperforms industry professionals on knowledge work benchmarks across 44 occupations. GPT-5.5 made multi-step agentic coding workflows viable without human supervision, GPT-5.6’s Sol tier leads the family on exactly that work, and GPT-6 has now landed on top of it without taking its place in Chat.

For most users the everyday choice inside ChatGPT is still GPT-5.6, Sol on paid plans and Luna on Free and Go, because that is what the conversation window serves. If you are calling the API or working in Codex, the cheaper answer is now the GPT-6 pair, GPT-6 Luna at $0.10 / $0.50 and GPT-6 Sol at $2 / $10, both half their GPT-5.6 counterparts and neither of them worse on the index. And when the job isn’t ChatGPT’s strength at all, our guide on when to use which AI shows which model to open instead. If you are building anything that will still be running in 12 months, plan for GPT-6 Astra at $10 / $50 per 1M tokens, which ChatGPT offers as GPT-6 Pro on Pro, Business and Enterprise.

FAQ

What is the latest ChatGPT model in 2026?

GPT-6. OpenAI launched GPT-6 Astra on September 3, 2026 and added GPT-6 Sol and GPT-6 Luna on September 22. In the ChatGPT conversation window, though, GPT-5.6 is still what answers you, Sol on paid plans and Luna on Free and Go, with Astra offered as the GPT-6 Pro option on Pro, Business and Enterprise. GPT-6 Sol and Luna run in ChatGPT Work, Codex and the API rather than in Chat. Sol scores 88.8% on Terminal-Bench 2.1 (91.9% in Sol Ultra mode) and all three GPT-5.6 tiers share a 1.05M-token context window.

What is the difference between GPT-5.4 and GPT-5.5?

GPT-5.5 outperforms GPT-5.4 on every frontier benchmark, most notably Terminal-Bench 2.0 (82.7% vs 75.1%), OSWorld-Verified (78.7% vs 75.0%), and SWE-Bench Pro (58.6% vs 57.7%). The trade-off is price, GPT-5.5 standard is $5 / $30 per 1M tokens, double GPT-5.4’s $2.50 / $15. GPT-5.5 Pro pricing matches GPT-5.4 Pro at $30 / $180.

What is the difference between GPT-5.3 and GPT-5.5?

GPT-5.5 adds native computer use, a 1M-token context window, and significantly stronger reasoning and agentic coding performance. GPT-5.3 Instant was dramatically cheaper at roughly $0.30 / $1.20, but it is no longer sold through the API, so the current budget equivalent is GPT-6 Luna at $0.10 / $0.50, half the GPT-5.6 Luna rate and below the legacy GPT-5.4-nano at $0.20 / $1.25 on both sides.

Is GPT-5.5 worth the price?

For agentic workflows, computer use, and professional coding, yes, though GPT-5.6 Sol costs less, $4 / $20 against GPT-5.5’s $5 / $30, and scores higher on agentic coding. GPT-6 Sol halves that again at $2 / $10 and edges GPT-5.6 Sol on the index, so in the API it is the better buy. For standard writing and chat, GPT-6 Luna at $0.10 / $0.50 covers the job for a fraction of the price. Use the cheap tier by default and reach for the flagship only when the task genuinely needs it.

Is GPT-6 out yet?

Yes, and it is now three models. GPT-6 Astra launched on September 3, 2026 at $10 / $50 per 1M tokens and reached ChatGPT the next day as the GPT-6 Pro option on Pro, Business and Enterprise. Plus plans get it in ChatGPT Work and Codex rather than in Chat, and Free and Go do not get it at all. GPT-6 Sol at $2 / $10 and GPT-6 Luna at $0.10 / $0.50 followed on September 22, 2026, in Work, Codex and the API, with Luna also in the desktop app for Free and Go.

Which ChatGPT models are being retired in 2026?

Fifteen legacy snapshots lost API access on July 23, 2026, including five Codex snapshots (gpt-5-codex, gpt-5.1-codex, gpt-5.1-codex-max, gpt-5.1-codex-mini, gpt-5.2-codex), both chat-latest aliases, both deep-research snapshots, and the original computer-use preview. A larger wave follows on October 23, 2026, covering GPT-3.5 Turbo, GPT-4, GPT-4 Turbo, o1, o3-mini, o4-mini, and gpt-image-1, then a third on December 11, 2026 takes the original GPT-5 snapshots, o3, and o3-pro. All three are API-only; the ChatGPT model picker is unaffected.

Which ChatGPT model is best for coding?

In Codex and the API, GPT-6 Sol is now the value pick for agentic, multi-step engineering at $2 / $10, with GPT-6 Astra above it for the hardest long-horizon work. Inside ChatGPT Chat the top pick is still GPT-5.6 Sol, with Extra High for the hardest problems (88.8% and 91.9% on Terminal-Bench 2.1). GPT-5.5 and GPT-5.5 Pro remain strong alternatives, and GPT-5.3-Codex is still a solid lower-cost option for large-codebase work with its 1M context.