| Result | Rank 1GPT-6.1 Sol Max | Rank 2GPT-6 Astra Low | Rank 3GPT-6 Astra High | Rank 4GPT-6.1 Sol Low | Rank 5GPT-6 Luna Low | Rank 6Composer 2.5 Default |
|---|---|---|---|---|---|---|
| Combined score | 88.04 | 85.84 | 80.89 | 79.72 | 41.50 | 35.54 |
| Judge rank | 1 | 2 | 3 | 4 | 5 | 6 |
| Judge score /100 | 92.50 | 87.50 | 84.00 | 81.00 | 39.00 | 35.00 |
| Judge breakdown | Judge breakdown
| Judge breakdown
| Judge breakdown
| Judge breakdown
| Judge breakdown
| Judge breakdown
|
| Time | 717.4 s | 47.28 s | 181.28 s | 45.95 s | 9.3 s | 72.57 s |
| Estimated USD | $0.451714 | N/A | N/A | $0.194894 | $0.0092994 | $0.071348 |
Reference estimates use original frozen rates; cost was excluded from this battle score; not billed USD.
Judge rank is derived from recorded rubric metrics across judged configurations; final rank uses the combined score.
Component points
| Result | Rank 1 GPT-6.1 Sol Max | Rank 2 GPT-6 Astra Low | Rank 3 GPT-6 Astra High | Rank 4 GPT-6.1 Sol Low | Rank 5 GPT-6 Luna Low | Rank 6 Composer 2.5 Default |
|---|---|---|---|---|---|---|
| Quality points | 87.6316 | 82.8947 | 79.5789 | 76.7368 | 36.9474 | 33.1579 |
| Test points | 0 | 0 | 0 | 0 | 0 | 0 |
| Speed points | 0.4062 | 2.9437 | 1.3088 | 2.9806 | 4.5569 | 2.3821 |
| Cost points | 0 | 0 | 0 | 0 | 0 | 0 |
Swipe, scroll or use Previous and More to compare configurations. Row labels stay visible.
Advanced results and scoring details
orchestrators · 2026-10-08T23:46:33.692Z · Prospective evaluation. Equal transport and native system conditions are unverified.
Applied weights: quality 94.7368%, tests 0%, speed 5.2632%, cost 0%.
Cost was omitted for every configuration in this cohort. Active weights were normalized uniformly. Displayed scores are rounded; pairwise outcomes use exact recorded fractions. Reference costs are estimates, not billed charges. Missing components follow the recorded assessment.
Pairwise results
- Composer 2.5 Default vs GPT-6 Astra High: Second configuration wins
- Composer 2.5 Default vs GPT-6 Luna Low: Second configuration wins
- Composer 2.5 Default vs GPT-6.1 Sol Low: Second configuration wins
- Composer 2.5 Default vs GPT-6 Astra Low: Second configuration wins
- Composer 2.5 Default vs GPT-6.1 Sol Max: Second configuration wins
- GPT-6 Astra High vs GPT-6 Luna Low: First configuration wins
- GPT-6 Astra High vs GPT-6.1 Sol Low: First configuration wins
- GPT-6 Astra High vs GPT-6 Astra Low: Second configuration wins
- GPT-6 Astra High vs GPT-6.1 Sol Max: Second configuration wins
- GPT-6 Luna Low vs GPT-6.1 Sol Low: Second configuration wins
- GPT-6 Luna Low vs GPT-6 Astra Low: Second configuration wins
- GPT-6 Luna Low vs GPT-6.1 Sol Max: Second configuration wins
- GPT-6.1 Sol Low vs GPT-6 Astra Low: Second configuration wins
- GPT-6.1 Sol Low vs GPT-6.1 Sol Max: Second configuration wins
- GPT-6 Astra Low vs GPT-6.1 Sol Max: Second configuration wins