Battle result
5 configurations · 28 battles
quality · 2026-10-07T05:51:36.615Z · Final combined score
Recovered original evaluation; this was not a new model run. Equal transport and native system conditions are unverified.
Applied weights: quality 90%, tests 0%, speed 5%, cost 5%.
| Result | Rank 1DeepSeek Flash · N/A | Rank 2GPT-6.1 Sol · Low | Rank 3Composer 2.5 · N/A | Rank 4GPT-6 Luna · Low | Rank 5Grok 4.7 · High |
|---|---|---|---|---|---|
| Combined score | 85.1385 | 84.1245 | 80.4989 | 79.8102 | 78.1299 |
| Effort | N/A | Low | N/A | Low | High |
| Quality points | 79.2 | 81 | 75.6 | 72 | 75.6 |
| Test points | 0 | 0 | 0 | 0 | 0 |
| Speed points | 3.0797 | 2.0109 | 3.0672 | 3.2831 | 1.7328 |
| Cost points | 2.8588 | 1.1136 | 1.8317 | 4.5271 | 0.7972 |
| Time | 37.411 s | 89.188 s | 37.808 s | 31.376 s | 113.133 s |
| Reference USD | $0.0074898 | $0.034898 | $0.0172975 | $0.0010446 | $0.052722 |
Swipe, scroll or use Previous and More to compare configurations. Row labels stay visible.
Displayed scores are rounded; pairwise outcomes use exact recorded fractions. Reference costs are estimates, not billed charges. Missing components follow the recorded assessment.
Pairwise results
- GPT-6.1 Sol · Low vs Composer 2.5 · N/A: First configuration wins
- GPT-6.1 Sol · Low vs DeepSeek Flash · N/A: Second configuration wins
- GPT-6.1 Sol · Low vs Grok 4.7 · High: First configuration wins
- GPT-6.1 Sol · Low vs GPT-6 Luna · Low: First configuration wins
- Composer 2.5 · N/A vs DeepSeek Flash · N/A: Second configuration wins
- Composer 2.5 · N/A vs Grok 4.7 · High: First configuration wins
- Composer 2.5 · N/A vs GPT-6 Luna · Low: First configuration wins
- DeepSeek Flash · N/A vs Grok 4.7 · High: First configuration wins
- DeepSeek Flash · N/A vs GPT-6 Luna · Low: First configuration wins
- Grok 4.7 · High vs GPT-6 Luna · Low: Second configuration wins