Battle result
5 configurations · 28 battles
quality · 2026-10-07T05:52:04.204Z · Final combined score
Recovered original evaluation; this was not a new model run. Equal transport and native system conditions are unverified.
Applied weights: quality 90%, tests 0%, speed 5%, cost 5%.
| Result | Rank 1GPT-6.1 Sol · Low | Rank 2Grok 4.7 · High | Rank 3DeepSeek Flash · N/A | Rank 4GPT-6 Luna · Low | Rank 5Composer 2.5 · N/A |
|---|---|---|---|---|---|
| Combined score | 84.5454 | 83.192 | 80.6474 | 80.4439 | 75.5273 |
| Effort | Low | High | N/A | Low | N/A |
| Quality points | 81 | 81 | 74.25 | 72 | 70.65 |
| Test points | 0 | 0 | 0 | 0 | 0 |
| Speed points | 2.4233 | 1.4499 | 3.3605 | 3.9057 | 3.0326 |
| Cost points | 1.1221 | 0.7421 | 3.0369 | 4.5382 | 1.8446 |
| Time | 63.798 s | 146.915 s | 29.272 s | 16.81 s | 38.924 s |
| Reference USD | $0.034558 | $0.057374 | $0.0064644 | $0.0010176 | $0.0171055 |
Swipe, scroll or use Previous and More to compare configurations. Row labels stay visible.
Displayed scores are rounded; pairwise outcomes use exact recorded fractions. Reference costs are estimates, not billed charges. Missing components follow the recorded assessment.
Pairwise results
- Grok 4.7 · High vs GPT-6 Luna · Low: First configuration wins
- Grok 4.7 · High vs Composer 2.5 · N/A: First configuration wins
- Grok 4.7 · High vs GPT-6.1 Sol · Low: Second configuration wins
- Grok 4.7 · High vs DeepSeek Flash · N/A: First configuration wins
- GPT-6 Luna · Low vs Composer 2.5 · N/A: First configuration wins
- GPT-6 Luna · Low vs GPT-6.1 Sol · Low: Second configuration wins
- GPT-6 Luna · Low vs DeepSeek Flash · N/A: Second configuration wins
- Composer 2.5 · N/A vs GPT-6.1 Sol · Low: Second configuration wins
- Composer 2.5 · N/A vs DeepSeek Flash · N/A: Second configuration wins
- GPT-6.1 Sol · Low vs DeepSeek Flash · N/A: First configuration wins