▸ compare
Anthropic: Claude Sonnet 4.6 vs Z.ai: GLM 5
Side-by-side benchmark comparison based on independent public data. Scores, pricing, context, and task breakdown.
▸ verdict
Anthropic: Claude Sonnet 4.6
89.3
higher score
vs
Z.ai: GLM 5
85.5
Anthropic: Claude Sonnet 4.6 leads with a higher composite benchmark score. It's the stronger choice for general-purpose tasks based on public benchmark data.
▸ score breakdown
| Task | Anthropic: Claude Sonnet 4.6 | Z.ai: GLM 5 | Δ |
|---|---|---|---|
| Overall | 89.3 | 85.5 | +3.8 |
| Coding | 56.2 | 65.2 | -9.0 |
| Reasoning | 67.0 | — | — |
| Math | — | — | — |
| Writing | — | — | — |
| JSON | — | — | — |
▸ specs & pricing
| Attribute | Anthropic: Claude Sonnet 4.6 | Z.ai: GLM 5 |
|---|---|---|
| Provider | Anthropic | Z Ai |
| Context window | 1M | 205K |
| Input $/M tokens | $3.00/M | $0.60/M |
| Output $/M tokens | $15.00/M | $1.92/M |
| Weights | proprietary | proprietary |
▸ frequently asked
Is Anthropic: Claude Sonnet 4.6 better than Z.ai: GLM 5?
Anthropic: Claude Sonnet 4.6 scores higher overall (89.3 vs 85.5) in the benchmark composite. The best choice depends on the specific use case and budget.
Which is cheaper: Anthropic: Claude Sonnet 4.6 or Z.ai: GLM 5?
Z.ai: GLM 5 is cheaper at $0.6/M input tokens vs $3/M for Anthropic: Claude Sonnet 4.6.
Which model is better for coding?
Z.ai: GLM 5 leads on coding with a score of 65.2 vs 56.2 for Anthropic: Claude Sonnet 4.6. This is based on SWE-Bench, Aider Polyglot, and LiveCodeBench data.