Battle result
5 configurations · 28 battles
quality · 2026-10-07T05:30:19.106Z · Final combined score
Recovered original evaluation; this was not a new model run. Equal transport and native system conditions are unverified.
Applied weights: quality 90%, tests 0%, speed 5%, cost 5%.
| Result | Rank 1GPT-6 Luna · Low | Rank 2Grok 4.7 · High | Rank 3GPT-6.1 Sol · Low | Rank 4DeepSeek Flash · N/A | Rank 5Composer 2.5 · N/A |
|---|---|---|---|---|---|
| Combined score | 87.2839 | 85.9755 | 85.3356 | 81.2653 | 75.8032 |
| Effort | Low | High | Low | N/A | N/A |
| Quality points | 79.2 | 84.6 | 81 | 74.25 | 70.65 |
| Test points | 0 | 0 | 0 | 0 | 0 |
| Speed points | 3.4897 | 0.8986 | 2.82 | 3.7025 | 3.1934 |
| Cost points | 4.5942 | 0.477 | 1.5156 | 3.3128 | 1.9598 |
| Time | 25.967 s | 273.871 s | 46.384 s | 21.027 s | 33.944 s |
| Reference USD | $0.0008832 | $0.094832 | $0.02299 | $0.0050928 | $0.0155125 |
Swipe, scroll or use Previous and More to compare configurations. Row labels stay visible.
Displayed scores are rounded; pairwise outcomes use exact recorded fractions. Reference costs are estimates, not billed charges. Missing components follow the recorded assessment.
Pairwise results
- Composer 2.5 · N/A vs Grok 4.7 · High: Second configuration wins
- Composer 2.5 · N/A vs DeepSeek Flash · N/A: Second configuration wins
- Composer 2.5 · N/A vs GPT-6 Luna · Low: Second configuration wins
- Composer 2.5 · N/A vs GPT-6.1 Sol · Low: Second configuration wins
- Grok 4.7 · High vs DeepSeek Flash · N/A: First configuration wins
- Grok 4.7 · High vs GPT-6 Luna · Low: Second configuration wins
- Grok 4.7 · High vs GPT-6.1 Sol · Low: First configuration wins
- DeepSeek Flash · N/A vs GPT-6 Luna · Low: Second configuration wins
- DeepSeek Flash · N/A vs GPT-6.1 Sol · Low: Second configuration wins
- GPT-6 Luna · Low vs GPT-6.1 Sol · Low: First configuration wins