Model · OpenAI
OpenAI Gpt Terra 5.6
1 appearance · meta-score 65.4/100 · last seen 2026-07-12 · scored under rubric v15
Pooled composite across 1 primary run: 61.5 — no interval, a single run proves nothing about spread provisional: n < 3, treat as unsettled
- $32.27Total cost
- 46,045,844Tokens
- 87.8%Cache hit
- 829Requests
- 0.00%Failure rate
- 1.00Tools/response
- 1,426,970Tokens per $
Agent role scores
| Agent role | Score / 10 |
|---|---|
| Sprint-Plan Agent | 7.51 |
| Scout Agent | 7.51 |
| Orchestrator | 7.09 |
| Architect | 6.89 |
| Test Agent | 6.89 |
| Code-Review Agent | 6.76 |
| Spec Agent | 6.72 |
| Pseudocode Agent | 6.56 |
| Develop Agent | 6.26 |
| Code Agent | 5.66 |
| Platform Agent | 5.17 |
Run history (primary model)
| Run | SoW | Date | Composite |
|---|---|---|---|
| Gpt Terra 5.6 | circles | Jul 2026 | 61.5 |
Every run on Favur Evals is scored by the same deterministic engine, on the same Statements of Work. The benchmark is self-funded — no vendor sponsorship, credits, or grants.