Favur Evals
favur.dev

Model · OpenAI

OpenAI Gpt Terra 5.6

1 appearance · meta-score 65.4/100 · last seen 2026-07-12 · scored under rubric v15

Pooled composite across 1 primary run: 61.5 — no interval, a single run proves nothing about spread provisional: n < 3, treat as unsettled

Agent role scores

Per-agent-role scores for OpenAI Gpt Terra 5.6
Agent roleScore / 10
Sprint-Plan Agent7.51
Scout Agent7.51
Orchestrator7.09
Architect6.89
Test Agent6.89
Code-Review Agent6.76
Spec Agent6.72
Pseudocode Agent6.56
Develop Agent6.26
Code Agent5.66
Platform Agent5.17

Run history (primary model)

Runs where OpenAI Gpt Terra 5.6 was the primary model, oldest first
RunSoWDateComposite
Gpt Terra 5.6circlesJul 202661.5

Every run on Favur Evals is scored by the same deterministic engine, on the same Statements of Work. The benchmark is self-funded — no vendor sponsorship, credits, or grants.