Vendor
OpenAI
2 model versions evaluated ·2 runs in the corpus · rubric v15 · OpenAI site ↗
Model versions
| Model | Meta-score | Appearances | Best at | Last seen |
|---|---|---|---|---|
| OpenAI Gpt Luna 5.6 | 64.7 | 1 | Sprint-Plan Agent | 2026-07-14 |
| OpenAI Gpt Terra 5.6 | 65.4 | 1 | Sprint-Plan Agent | 2026-07-12 |
Runs (primary model)
| Run | SoW | Date | Composite |
|---|---|---|---|
| Gpt Terra 5.6 | circles | Jul 2026 | 61.5 |
| Gpt Luna 5.6 | circles | Jul 2026 | 61.1 |
Every run on Favur Evals is scored by the same deterministic engine, on the same Statements of Work. The benchmark is self-funded — no vendor sponsorship, credits, or grants.