WritingmateWritingmate

AI model reference · 2026

GPT-5.4

OpenAI: GPT-5.4 logoOpenAIModel details, evidence, and API facts

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system.

Text generation
Image input
Reasoning
Provider endpoint accepts tool parameters
Context window
1,050,000 tokens
Maximum output
128,000 tokens
Base provider API input
$2.50 / 1M tokens
Starting provider API rate; context tiers apply
Published evidence
2 benchmarks, 2 Arena results

Overview

About GPT-5.4

A concise catalog overview. Technical limits and published evaluation evidence are listed separately below.

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system.

text
multimodal
reasoning

Published evaluations

Benchmark results

Only results mapped to this exact model and backed by a named source appear here. Different evaluation protocols are not treated as interchangeable.

Source-backed benchmark results for GPT-5.4
BenchmarkScoreEvaluation detailsSource
DeepSWE 1.1A long-horizon software-engineering benchmark with 113 original tasks graded by hand-written tests.52.0%±2Pass@1Version: DeepSWE 1.1 · Split: 113 tasks across 91 repositories · Harness: mini-swe-agent · Reasoning effort: xhighProtocol note: All models use the same harness. Cost, output-token, and step counts are retained per result as protocol context.; Confidence interval: ±2Methodology DataCurveBenchmark owner
LiveBenchA contamination-resistant benchmark refreshed on a fixed release cadence. Scores from different LiveBench releases must never be compared as the same protocol.78.0%Mean of category averagesVersion: LiveBench 2026-06-25 · Split: 2026-06-25 release, 23 tasks across seven categories · Reasoning effort: xhighProtocol note: Overall is the mean of category averages. This protocol is not comparable with the 2026-01-08 release.Methodology LiveBenchBenchmark owner

Human preference

Arena results

Arena scores come from blind human preference votes. They are reported separately from task benchmarks and are not used as substitutes for missing benchmark results.

Arena results for GPT-5.4
CategoryScoreRankVotesStatus
Arena TextBlind human preference score for text responses.1,465±4#4263,097Published
Arena WebDevBlind human preference score for code and web development.1,393±16#641,461Published

Technical reference

Specifications and API facts

Technical limits and reference provider prices come from the current catalog unless a reviewed external source is listed. Missing values stay marked as unavailable.

Model
Provider
OpenAI
Release date
Not available in reviewed sources
Catalog added
Mar 5, 2026
Context window
1,050,000 tokens
Maximum output
128,000 tokens
Knowledge cutoff
Not available in reviewed sources
License
Not available in reviewed sources
Capabilities and API
Input types
Text, Image, File
Output types
Text
Reasoning
Supported
Provider endpoint tool parameters
Accepts tool parameters
Base provider API input
$2.50 / 1M tokens
Base provider cached input
$0.25 / 1M tokens
Base provider API output
$15 / 1M tokens
Provider tier at ≥ 272,000 prompt tokens
Input $5.00 / 1M tokensOutput $22.5 / 1M tokensCached input $0.50 / 1M tokens

Provider token rates are context-dependent. The base rates apply below the listed prompt thresholds; the matching tier applies at or above each threshold.

These are reference provider API rates, not Writingmate checkout charges. Access in Writingmate follows the allowances of your Writingmate plan.

Provenance

Sources and update status

Source links are attached to the facts and evaluations they support. Catalog-only values are not presented as independently verified claims.

Evidence last updated Aug 9, 2026.

Catalog record updated Mar 5, 2026.

Catalog-added dates describe when a model entered the catalog, not necessarily its public release date.

Try GPT-5.4 in Writingmate

Use this model, then compare its response with other AI models in the same workspace.