by OpenAI · GPT-5 family · best for escalation tier for in-flight GPT-5.4 stacks
GPT-5.4 Pro is the deep-reasoning variant of GPT-5.4, released 2026-03-05 — the same base model with reasoning effort fixed to medium/high/xhigh and a tool surface tuned for analytical work. With GPT-5.5 Pro now shipping at the identical $30/$180 price point but on a newer base, GPT-5.4 Pro occupies a narrow niche: workloads that need the GPT-5.4 reasoning ceiling and its specific tool surface (apply_patch, computer use, tool search) without the disruption of switching base models. The one-sentence buyer's take: a defensible bridge SKU for in-flight GPT-5.4 stacks, but most new spend at this price should target GPT-5.5 Pro.
| Benchmark | Score | Source |
|---|---|---|
| GPQA Diamond | 93% | developers.openai.com 2026-03-05T00:00:00.000Z |
Six personas, six verdicts — the same panel that reviews every product on TopReviewed.
“Boxed in by a same-priced successor — defensible only as continuity for in-flight GPT-5.4 stacks, not for new spend.”
GPT-5.4 Pro is in an awkward spot. It costs the same as GPT-5.5 Pro but uses an older base with a four-month-older knowledge cutoff. The argument for it is continuity — teams already standardized on GPT-5.4 can adopt Pro without re-validating a new base, and the tool surface differs (apply_patch is here, not on 5.5 Pro). Strategic recommendation: most new builds target GPT-5.5 Pro; GPT-5.4 Pro is a sensible bridge for in-flight migrations. Expect share to decline over two quarters as migration completes.
“A transitional SKU with one real wedge — the GPT-5.4 tool surface at the Pro tier — that GPT-5.5 Pro can't match.”
Strategically this is a transition product. Its only durable differentiation against GPT-5.5 Pro is the tool surface (apply_patch, computer use, tool search), which matters to a narrow set of agentic-research workflows. On every other axis — knowledge cutoff, base capability, market momentum — GPT-5.5 Pro is ahead at the same price. Market timing works against it: launching three weeks before GPT-5.5 Pro meant a short window as the deep-reasoning default. Expect it to serve continuity needs and then fade.
“Financially identical to GPT-5.5 Pro — the only question is whether the older base is acceptable, and for new spend it usually isn't.”
Pricing is identical to GPT-5.5 Pro: $30/$180, no cached-input discount, Batch at 50%. The financial question is purely whether the older base is acceptable for the workload. For teams mid-migration, the cost of switching base models can outweigh the marginal capability gain — Pro 5.4 buys continuity. For new spend at this price, GPT-5.5 Pro is the default recommendation. Value-per-dollar is poor given a same-priced, more-capable alternative exists.
“A clean escalation for GPT-5.4 shops — same API shape, deeper reasoning, apply_patch intact — but no code interpreter or hosted shell.”
For developers already shipping on GPT-5.4, Pro is a clean escalation — same API shape, same tool semantics, deeper reasoning, apply_patch supported (which matters for code-edit agents). The choice against GPT-5.5 Pro comes down to two things: do you need apply_patch/tool search (Pro 5.4 yes, Pro 5.5 no), and do you need the newer knowledge cutoff (Pro 5.5 yes). The lack of code interpreter and hosted shell here limits active sandbox workflows. Tool calling and structured outputs are stable; background mode is required for the longest runs.
“Surfaces as a 'deeper analysis' tier — careful answers, slow latency, and rarely visibly better than GPT-5.5 Pro on the same task.”
End users don't pick this directly — it surfaces inside ChatGPT and vertical apps as a deeper-analysis tier. Quality is high but not visibly better than GPT-5.5 Pro on most tasks, and slower latency is the main visible cost. For research-grade questions the answers are markedly more careful than GPT-5.4 base; for casual chat, indistinguishable. Refusal patterns mirror the base model.
“OpenAI priced an older Pro identically to its newer Pro — the honest read is continuity revenue, not a distinct product.”
The adversarial read: GPT-5.4 Pro and GPT-5.5 Pro at the same $30/$180 is a pricing choice that mostly benefits OpenAI's migration narrative, not buyers. The one genuine differentiator — the GPT-5.4 tool surface — is real but narrow, and the published benchmark trail is thin (approximate GPQA and SWE-bench, nothing else). Reasoning-token billing with no cache discount makes worst-case cost ugly. This is a competent model whose main reason to exist is to not disrupt teams that already standardized on GPT-5.4; for almost everyone else, GPT-5.5 Pro is strictly better at the same price.
The full research notes behind this review — verified against primary sources.
Same undisclosed architecture as GPT-5.4 base — no published parameter, layer, or dense/MoE detail, all null. The distinction from base is operational: reasoning effort is pinned to the upper tiers and per-request compute budgets are larger. Unified reasoning model, text + image input, text-only output, o200k_base tokenizer, 1.05M context with a 272K break-point.
GPT-5.4 Pro is the reasoning ceiling on the GPT-5.4 base, justifying its 9.3 reasoning and 9.2 math scores — consistently 1–3 points above GPT-5.4 base on reasoning-heavy evals (GPQA Diamond ~93%) at a real latency cost. Coding (9.2) is strong for careful one-shot solutions and multi-step investigations; agentic (8.7) is capped because, unlike base, Pro is meant for deliberate answers rather than active sandbox manipulation (no code interpreter or hosted shell here). Long-context (8.5) benefits from deeper deliberation. Instruction-following (9.2) and safety calibration (8.9) are strong. Vision (8.4) and document/OCR (8.2) match base. Function calling (8.8) works on the supported tool set. Real-time data (6.8) is web-search-dependent.
| Benchmark | Score | vs Base (GPT-5.4) | vs Top Competitor (GPT-5.5 Pro) | Source |
|---|---|---|---|---|
| GPQA Diamond | ~93% | up from 92.8% base | -1pp vs 5.5 Pro (~94%) | developers.openai.com |
| SWE-bench Verified | ~83% | up from base | trailing 5.5 Pro (~89%) | developers.openai.com |
Pro variants ship aggregate numbers; GPT-5.4 Pro runs ~1–3 points above GPT-5.4 base on reasoning-heavy evals at a latency cost. HLE, AIME, MMLU-Pro, LiveCodeBench, Tau-bench, and LMArena have no separately-published GPT-5.4 Pro figure as of 2026-05-28 and are recorded null. research_confidence is medium.
Slower than GPT-5.4 base by design — Pro always runs extended reasoning, so requests take meaningfully longer. OpenAI does not publish a steady-state tokens/sec or TTFT figure for the Pro SKU (recorded null). Streaming has restrictions; the longest requests need background mode. latency_tier is slow.
| Surface | Cost | Notes |
|---|---|---|
| API input | $30.00 / 1M tok | no cached-input discount |
| API output | $180.00 / 1M tok | |
| Batch (in/out) | $15.00 / $90.00 | 50% off, 24h SLA |
| Direct UI | $200/mo (Pro) | included as deep-thinking option |
| Free tier | none | |
| Streaming | restricted | use background mode for long requests |
Reasoning-token note: Pro always reasons at upper-tier effort, so hidden reasoning tokens (billed as output) dominate cost. No cached-input discount; Batch is the only cost lever.
API-only via the Responses API; not open-weights, license Proprietary, not self-hostable. Cloud-managed via Azure OpenAI and Azure AI Foundry; OpenRouter proxies it. Data residency covers US and EU. The differentiating deployment factor versus GPT-5.5 Pro is the tool surface — apply_patch, computer use, and tool search are available here but not on GPT-5.5 Pro.
Governed by OpenAI's Preparedness Framework. No training on API inputs by default; opt-out and zero-retention available for enterprise. Compliance covers SOC2, GDPR, CCPA, and HIPAA (BAA). Content moderation is built in. Refusal patterns mirror GPT-5.4 base.
Same first-party SDKs (Python, TypeScript, Java, Go, .NET) and OpenAI Agents SDK as base. Surfaces in ChatGPT Pro/Business as a deeper-analysis mode and in Deep Research. Popularity tier: niche, concentrated among in-flight GPT-5.4 stacks.
Only two reasons: you need the GPT-5.4 tool surface (apply_patch, computer use, tool search), or you want to avoid re-validating a new base model mid-migration.
Reasoning effort is pinned to medium/high/xhigh, there is no cache discount, streaming is restricted, and code interpreter/hosted shell are dropped.
Batch only (50% off). No cached-input discount exists on Pro.
Not on the deprecation list as of 2026-05-28, but expect declining relevance as GPT-5.5 Pro takes new spend.
No, not by API default; enterprise opt-out and zero-retention exist.
same $30/$180 price, newer December-2025 base, but a smaller tool surface (no apply_patch/tool search); the default for new spend.
same model at default reasoning effort, 12x cheaper, with the full tool surface plus code interpreter and hosted shell.
peer escalation tier, often cheaper per answer with stronger long-form quality.
Primary references used to verify this review.
Does not train on API inputs by default
Last verified 2026-05-27