Favur Evals
favur.dev

Model · Z.ai

Z.ai Glm 5.2

3 appearances · meta-score 69.9/100 · last seen 2026-07-17 · scored under rubric v15

Pooled composite across 3 primary runs: 54.3 ± 18.3 (95% CI)

Agent role scores

Per-agent-role scores for Z.ai Glm 5.2
Agent roleScore / 10
Sprint-Review Agent8.57
Orchestrator7.64
Scout Agent7.60
Develop Agent7.56
Build Agent7.19
Sprint-Plan Agent6.96
Code-Review Agent6.93
Code Agent6.42
Test Agent6.25
Pseudocode Agent5.39

Run history (primary model)

Runs where Z.ai Glm 5.2 was the primary model, oldest first
RunSoWDateComposite
Glm 5.2circlesJun 202646.9
Glm 5.2circlesJul 202654.5
Glm 5.2circlesJul 202661.6

Every run on Favur Evals is scored by the same deterministic engine, on the same Statements of Work. The benchmark is self-funded — no vendor sponsorship, credits, or grants.