AI model reference · 2026
GLM 5.1
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks.
- Context window
- 204,800 tokens
- Maximum output
- 131,072 tokens
- Base provider API input
- $1.40 / 1M tokens
- Base provider API rate
- Published evidence
- 1 benchmark, 2 Arena results
Overview
About GLM 5.1
A concise catalog overview. Technical limits and published evaluation evidence are listed separately below.
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks.
Published evaluations
Benchmark results
Only results mapped to this exact model and backed by a named source appear here. Different evaluation protocols are not treated as interchangeable.
| Benchmark | Score | Evaluation details | Source |
|---|---|---|---|
| Terminal-Bench 2.1Version 2.1 of the benchmark for completing realistic tasks in terminal environments. Harness, resource limits, and attempt count are part of the protocol. | 58.7%±1.2 | Mean task successVersion: Terminal-Bench 2.1, Harbor dataset revision 6 · Split: 89 revised tasks · Harness: Claude Code · Attempts: Five attempts per task · Reasoning effort: maxProtocol note: Every result retains its agent. Rows with different agents are not model-only comparisons.; Agent: Claude Code; Run date: 2026-05-01; Confidence interval: ±1.2Methodology | Terminal-BenchBenchmark owner |
Human preference
Arena results
Arena scores come from blind human preference votes. They are reported separately from task benchmarks and are not used as substitutes for missing benchmark results.
| Category | Score | Rank | Votes | Status |
|---|---|---|---|---|
| Arena TextBlind human preference score for text responses. | 1,468±4 | #36 | 36,716 | Published |
| Arena WebDevBlind human preference score for code and web development. | 1,515±8 | #26 | 7,853 | Published |
Technical reference
Specifications and API facts
Technical limits and reference provider prices come from the current catalog unless a reviewed external source is listed. Missing values stay marked as unavailable.
- Provider
- Z.AI
- Release date
- Not available in reviewed sources
- Catalog added
- Apr 7, 2026
- Context window
- 204,800 tokens
- Maximum output
- 131,072 tokens
- Knowledge cutoff
- Not available in reviewed sources
- License
- Not available in reviewed sources
- Input types
- Text
- Output types
- Text
- Reasoning
- Supported
- Provider endpoint tool parameters
- Accepts tool parameters
- Base provider API input
- $1.40 / 1M tokens
- Base provider cached input
- $0.26 / 1M tokens
- Base provider API output
- $4.40 / 1M tokens
These are reference provider API rates, not Writingmate checkout charges. Access in Writingmate follows the allowances of your Writingmate plan.
Provenance
Sources and update status
Source links are attached to the facts and evaluations they support. Catalog-only values are not presented as independently verified claims.
- Arena TextArena AI · Retrieved Aug 9, 2026
- Arena WebDevArena AI · Retrieved Aug 9, 2026
- terminal-bench@2.1 LeaderboardTerminal-Bench · Retrieved Aug 9, 2026
- OpenRouter model catalogOpenRouter
Evidence last updated Aug 9, 2026.
Catalog record updated Apr 7, 2026.
Catalog-added dates describe when a model entered the catalog, not necessarily its public release date.
Try GLM 5.1 in Writingmate
Use this model, then compare its response with other AI models in the same workspace.