AI model reference · 2026
GPT-5.2-Codex
GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows.
- Context window
- 400,000 tokens
- Maximum output
- 128,000 tokens
- Base provider API input
- $1.75 / 1M tokens
- Base provider API rate
- Published evidence
- 2 benchmarks, 1 Arena result
Overview
About GPT-5.2-Codex
A concise catalog overview. Technical limits and published evaluation evidence are listed separately below.
GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows.
Published evaluations
Benchmark results
Only results mapped to this exact model and backed by a named source appear here. Different evaluation protocols are not treated as interchangeable.
| Benchmark | Score | Evaluation details | Source |
|---|---|---|---|
| SWE-Bench VerifiedA human-validated subset of real GitHub issues used to measure whether a coding agent can produce repository patches that resolve the associated tests. | 72.8% | ResolvedVersion: SWE-Bench Verified 500 · Split: Verified, 500 tasks · Harness: mini-SWE-agent / version 2.0.0 · Attempts: One attemptProtocol note: Harness version: 2.0.0; Run date: 2026-02-19Methodology | SWE-benchBenchmark owner |
| LiveBenchA contamination-resistant benchmark refreshed on a fixed release cadence. Scores from different LiveBench releases must never be compared as the same protocol. | 74.0% | Mean of category averagesVersion: LiveBench 2026-06-25 · Split: 2026-06-25 release, 23 tasks across seven categoriesProtocol note: Overall is the mean of category averages. This protocol is not comparable with the 2026-01-08 release.Methodology | LiveBenchBenchmark owner |
Human preference
Arena results
Arena scores come from blind human preference votes. They are reported separately from task benchmarks and are not used as substitutes for missing benchmark results.
| Category | Score | Rank | Votes | Status |
|---|---|---|---|---|
| Arena WebDevBlind human preference score for code and web development. | 1,338±9 | #91 | 6,424 | Published |
Technical reference
Specifications and API facts
Technical limits and reference provider prices come from the current catalog unless a reviewed external source is listed. Missing values stay marked as unavailable.
- Provider
- OpenAI
- Release date
- Not available in reviewed sources
- Catalog added
- Jan 14, 2026
- Context window
- 400,000 tokens
- Maximum output
- 128,000 tokens
- Knowledge cutoff
- Not available in reviewed sources
- License
- Not available in reviewed sources
- Input types
- Text, Image
- Output types
- Text
- Reasoning
- Supported
- Provider endpoint tool parameters
- Accepts tool parameters
- Base provider API input
- $1.75 / 1M tokens
- Base provider cached input
- $0.175 / 1M tokens
- Base provider API output
- $14 / 1M tokens
These are reference provider API rates, not Writingmate checkout charges. Access in Writingmate follows the allowances of your Writingmate plan.
Provenance
Sources and update status
Source links are attached to the facts and evaluations they support. Catalog-only values are not presented as independently verified claims.
- Arena WebDevArena AI · Retrieved Aug 25, 2026
- LiveBench 2026-06-25 LeaderboardLiveBench · Retrieved Aug 9, 2026
- SWE-bench Verified Bash Only LeaderboardSWE-bench · Retrieved Aug 9, 2026
- OpenRouter model catalogOpenRouter
Evidence last updated Aug 25, 2026.
Catalog record updated Jan 14, 2026.
Catalog-added dates describe when a model entered the catalog, not necessarily its public release date.
Try GPT-5.2-Codex in Writingmate
Use this model, then compare its response with other AI models in the same workspace.