Model · Qwen
Qwen Flash qwen3.7
1 appearance · meta-score 70.0/100 · last seen 2026-08-10 · scored under rubric v15
Pooled composite across 1 primary run: 76.3 - no interval, a single run proves nothing about spread provisional: n < 3, treat as unsettled
- $2.66Total cost
- 68,128,133Tokens
- 72.1%Cache hit
- 1,036Requests
- 1.06%Failure rate
- 1.08Tools/response
- 25,633,736Tokens per $
Agent role scores
| Agent role | Score / 10 |
|---|---|
| Sprint-Review Agent | 8.88 |
| Code-Review Agent | 8.20 |
| Build Agent | 7.76 |
| Develop Agent | 7.59 |
| Orchestrator | 7.56 |
| Scout Agent | 6.54 |
| Test Agent | 6.47 |
| Sprint-Plan Agent | 6.19 |
| Pseudocode Agent | 5.62 |
| Code Agent | 5.28 |
Run history (primary model)
| Run | SoW | Date | Composite |
|---|---|---|---|
| Flash qwen3.7 | circles | Aug 2026 | 76.3 |
Every run on Favur Evals is scored by the same deterministic engine, on the same Statements of Work. The benchmark is self-funded - no vendor sponsorship, credits, or grants.