01 / decision
Decision snapshot
Each figure sits next to the middle of the field, so it can be read as dear or cheap, wide or narrow, rather than floating on its own. Compared against the 370 models tracked here, not against an absolute standard.
Capability
55.4/100
#22 of 35 with published scores
Input / 1M tokens
$0.43
field median $0.40
Output / 1M tokens
$1.75
field median $1.68
Context
205K
field median 262K
Read the coverage before the score. That figure rests on 20% of the weighting: the rest of the categories have no published evidence for this model, and are not counted rather than counted as zero.
02 / overview
What GLM 4.6 is, and when to reach for it
Four questions, answered separately, because somebody arrives at one of them rather than at the top of the page.
This model has no written explanation yet.
The page is created the moment a model appears in a provider's catalogue, and the writing follows. Until then the measured sections below are the whole page.
What counts as evidenceFollows GLM 4.5V. Superseded by GLM 4.6V. See the whole line.
03 / evidence
How much of this is verified
Split by category, so a strong number never hides a thin evidence base. Verified means it was read on the benchmark's own published results; a provider's claim about its own model is shown and labelled rather than dropped.
Published rows
1
of 19 benchmarks tracked
Independently verified
1
0 from the provider
Weighting covered
20%
the rest is not measured, not zero
Agentic
0/3
Not measured
Coding
1/3
Verified
Reasoning
0/3
Not measured
Multimodal
0/3
Not measured
Knowledge
0/2
Not measured
Multilingual
0/2
Not measured
Instruction following
0/1
Not measured
Maths
0/2
Not measured
04 / ledger
Benchmark ledger
Every published row, grouped by category, each compared with the best published score on the same benchmark. 'Is 64% good' is a question nobody can answer; '26 points behind the leader' is one anybody can.
Coding1 row
| Benchmark | Score | Best published | Weight | Evidence |
|---|---|---|---|---|
| SWE-bench VerifiedReal GitHub issues, resolved and tested. | 55.4% | Gemini 3 Flash Preview75.8% · 20.4 behind | 40% | Verified |
05 / capability
Capability shape
Where this model is strong, and against how many peers. Ranks are against models with evidence in that category, not against every tracked model: ranking against models nobody tested would rank who published, not who is better.
| Category | Score | Rank | Weight | Benchmarks | Evidence |
|---|---|---|---|---|---|
| Agentic | Not measured | Not ranked | 22% | 0 of 3 | Not measured |
| Coding | 55.4 | #22 of 3538th percentile | 20% | 1 of 3 | Verified |
| Reasoning | Not measured | Not ranked | 17% | 0 of 3 | Not measured |
| Multimodal | Not measured | Not ranked | 12% | 0 of 3 | Not measured |
| Knowledge | Not measured | Not ranked | 12% | 0 of 2 | Not measured |
| Multilingual | Not measured | Not ranked | 7% | 0 of 2 | Not measured |
| Instruction following | Not measured | Not ranked | 5% | 0 of 1 | Not measured |
| Maths | Not measured | Not ranked | 5% | 0 of 2 | Not measured |
Category weights are WriteWorks' own and they are a judgement: weighted for what a team buys a model to do. Agentic and coding lead because that is where the work is. Published here rather than hidden, and set out in full on the methodology page.
06 / cost
What it costs
List API rates as last read, each with the date it was confirmed, plus every change recorded since tracking began.
| Charge | Price | Unit | Read on |
|---|---|---|---|
| Input | $0.43 | per 1M tokens | 2026-09-30 |
| Output | $1.75 | per 1M tokens | 2026-09-30 |
| Workload | Input / month | Output / month | Cost |
|---|---|---|---|
| A small product team | 20M tokens | 5M tokens | $17.35 |
| A busy support assistant | 200M tokens | 40M tokens | $156.00 |
| A document pipeline | 1000M tokens | 100M tokens | $605.00 |
List API rates, no caching and no batch discount, which both providers offer and which change the answer a great deal. Treat these as the ceiling, not the bill.
07 / specs
Specifications
As listed by the provider's own catalogue and re-read every few hours. Anything absent is absent there too.
| Context window | 204,800 tokens |
|---|---|
| Maximum output | 16,384 tokens |
| Modalities | text |
| Released | 30/09/2025 |
| Status | Current |
08 / lineage
Lineage
What this model replaced, what replaced it, and what else its provider has in the field.
Also from Z.ai
- GLM 5.3 Prime23/09/2026
- GLM 5.3 FlashX18/09/2026
- GLM Flash Latest27/08/2026
- GLM 5.3 Flash26/08/2026
- GLM Latest19/08/2026
- GLM 5.318/08/2026
- GLM 5.216/06/2026
- GLM 5.107/04/2026
- GLM 5V Turbo01/04/2026
- GLM 5 Turbo15/03/2026
- GLM 511/02/2026
- GLM 4.7 Flash19/01/2026
09 / line
The line
Every model in this family in release order, so a page from eight months ago says in one glance that two newer ones exist.
Came before
← GLM 4.5VCame after
GLM 4.6V→- 01GLM 4.525/07/2025
- 02GLM 4.5 Air25/07/2025
- 03GLM 4.5V11/08/2025
- 04GLM 4.630/09/2025
- 05GLM 4.6V08/12/2025
- 06GLM 4.722/12/2025
- 07GLM 4.7 Flash19/01/2026
- 08GLM 511/02/2026
- 09GLM 5 Turbo15/03/2026
- 10GLM 5V Turbo01/04/2026
- 11GLM 5.107/04/2026
- 12GLM 5.216/06/2026
- 13GLM 5.318/08/2026
- 14GLM Latest19/08/2026
- 15GLM 5.3 Flash26/08/2026
- 16GLM 5.3 FlashX18/09/2026
- 17GLM 5.3 Prime23/09/2026
Ordered by release date and worked out from the naming, so a new member slots in as soon as its page exists. A retired model keeps its page and its place in the line.
10 / notes
WriteWorks notes
What this model changes for a brand trying to be cited in AI answers, and every change logged since it launched.
Change log
Every release, price and feature change for GLM 4.6, newest first. They also appear on the Z.ai page.
Nothing published here yet. Changes appear within hours of a provider announcing them.
11 / questions
Questions
The things people ask about this model, answered from what is on this page rather than from anywhere else.
- What does GLM 4.6 cost?
- $0.43 per million input tokens and $1.75 per million output tokens, as last read from the provider. The cost section works that into a monthly figure.
- How current is this page?
- The catalogue behind it is re-read every three hours, and the stamp at the top says when it last confirmed. A re-check that finds nothing changed updates that stamp and deliberately does not touch the page’s modified date.
- Why are some sections empty?
- Because nothing has been published that can be linked to. An empty section is better than a number you cannot check. The methodology sets out what counts.
Is GLM 4.6 recommending you?
Models change what gets cited. WriteWorks tracks whether 10+ AI platforms, from ChatGPT and Claude to Gemini and Perplexity, name your brand or your competitors, and shows you the content gaps to close.