live
weekly refresh
basedagi.org
▸ compare

Anthropic: Claude Sonnet 4.6 vs Z.ai: GLM 5

Side-by-side benchmark comparison based on independent public data. Scores, pricing, context, and task breakdown.

▸ verdict
Anthropic: Claude Sonnet 4.6
89.3
higher score
vs
Z.ai: GLM 5
85.5
Anthropic: Claude Sonnet 4.6 leads with a higher composite benchmark score. It's the stronger choice for general-purpose tasks based on public benchmark data.
▸ score breakdown
TaskAnthropic: Claude Sonnet 4.6Z.ai: GLM 5Δ
Overall89.385.5+3.8
Coding56.265.2-9.0
Reasoning67.0
Math
Writing
JSON
▸ specs & pricing
AttributeAnthropic: Claude Sonnet 4.6Z.ai: GLM 5
ProviderAnthropicZ Ai
Context window1M205K
Input $/M tokens$3.00/M$0.60/M
Output $/M tokens$15.00/M$1.92/M
Weightsproprietaryproprietary
▸ frequently asked

Is Anthropic: Claude Sonnet 4.6 better than Z.ai: GLM 5?

Anthropic: Claude Sonnet 4.6 scores higher overall (89.3 vs 85.5) in the benchmark composite. The best choice depends on the specific use case and budget.

Which is cheaper: Anthropic: Claude Sonnet 4.6 or Z.ai: GLM 5?

Z.ai: GLM 5 is cheaper at $0.6/M input tokens vs $3/M for Anthropic: Claude Sonnet 4.6.

Which model is better for coding?

Z.ai: GLM 5 leads on coding with a score of 65.2 vs 56.2 for Anthropic: Claude Sonnet 4.6. This is based on SWE-Bench, Aider Polyglot, and LiveCodeBench data.