MidMarketBench
Rank 7completeObserved OpenRouter runProvisional tie

GLM 5.2

Model and run

Provider
Z.ai
Version
glm-5.2
Release
Not published
Context
1,048,576 tokens
Weights
Not published
Run status
complete
Candidate samples
2/2
Scored samples
2/2
Overall
88.2
Observed range
86.6 to 89.7
Provider-reported attributed cost
$0.261
Candidate attempts
$0.018
Judge attempts
$0.243
Median candidate latency
2.3 minutes

Model source: OpenRouter catalogue

Attributed costs include every OpenRouter-reported candidate and judge attempt for this model, including retries. The run total is authoritative for settled key spend.

OpenRouter provenance

Requested model and returned route.

Requested model
z-ai/glm-5.2
Routed provider
StreamLake
Endpoint tag
streamlake/fp8
Returned model
z-ai/glm-5.2
Quantisation
FP8
Pinned route price
streamlake/fp8: $0.274 input / $0.860 output per 1M tokens

Run 2026-07-18-final / 2026-07-18 / Closed-book

Dimension shape

Where the observed score comes from.

Overall combines deterministic task checks with blinded, calibrated, cross-family judgements of the IC note.

Grounding85.9
Commercial judgement95.8
Scepticism88.4
Numerical sanity90.7
Risk discovery83.2
Question generation87.3
European context70.0
Output usefulness90.4

Complete field

Position among ranked two-sample runs.

6498

Interpretation limit

One case, 2 scored samples.

Directional mini benchmark: one fresh synthetic case and two samples per model. It is not a universal ranking of model intelligence.