Battle result
5 configurations · 28 battles
quality · 2026-10-07T05:38:00.895Z · Final combined score
Recovered original evaluation; this was not a new model run. Equal transport and native system conditions are unverified.
Applied weights: quality 90%, tests 0%, speed 5%, cost 5%.
| Result | Rank 1GPT-6.1 Sol · Low | Rank 2Grok 4.7 · High | Rank 3GPT-6 Luna · Low | Rank 4DeepSeek Flash · N/A | Rank 5Composer 2.5 · N/A |
|---|---|---|---|---|---|
| Combined score | 84.3127 | 82.8872 | 79.9598 | 69.4765 | 66.526 |
| Effort | Low | High | Low | N/A | N/A |
| Quality points | 81 | 81 | 72 | 61.65 | 61.65 |
| Test points | 0 | 0 | 0 | 0 | 0 |
| Speed points | 2.1246 | 1.2078 | 3.3995 | 4.0653 | 2.9853 |
| Cost points | 1.1881 | 0.6794 | 4.5602 | 3.7612 | 1.8907 |
| Time | 81.202 s | 188.38 s | 28.247 s | 13.795 s | 40.492 s |
| Reference USD | $0.032084 | $0.063592 | $0.0009644 | $0.0032937 | $0.0164455 |
Swipe, scroll or use Previous and More to compare configurations. Row labels stay visible.
Displayed scores are rounded; pairwise outcomes use exact recorded fractions. Reference costs are estimates, not billed charges. Missing components follow the recorded assessment.
Pairwise results
- Grok 4.7 · High vs GPT-6.1 Sol · Low: Second configuration wins
- Grok 4.7 · High vs Composer 2.5 · N/A: First configuration wins
- Grok 4.7 · High vs GPT-6 Luna · Low: First configuration wins
- Grok 4.7 · High vs DeepSeek Flash · N/A: First configuration wins
- GPT-6.1 Sol · Low vs Composer 2.5 · N/A: First configuration wins
- GPT-6.1 Sol · Low vs GPT-6 Luna · Low: First configuration wins
- GPT-6.1 Sol · Low vs DeepSeek Flash · N/A: First configuration wins
- Composer 2.5 · N/A vs GPT-6 Luna · Low: Second configuration wins
- Composer 2.5 · N/A vs DeepSeek Flash · N/A: Second configuration wins
- GPT-6 Luna · Low vs DeepSeek Flash · N/A: First configuration wins