| Result | Rank 1GPT-6.1 Sol Low | Rank 2GPT-6.1 Sol Medium | Rank 3GPT-6 Astra Medium | Rank 4GPT-6.1 Sol Extra High | Rank 5GPT-6 Luna Low | Rank 6Composer 2.5 Default |
|---|---|---|---|---|---|---|
| Combined score | 85.67 | 85.46 | 84.67 | 83.43 | 80.43 | 73.66 |
| Judge | ||||||
| Judge score /100 | 90.00 | 90.00 | 90.00 | 90.00 | 79.50 | 76.50 |
| Correctness /10 | 9 | 9 | 9 | 9 | 8 | 7 |
| Coverage /10 | 9 | 9 | 9 | 9 | 7 | 9 |
| Evidence /10 | 9 | 9 | 9 | 9 | 9 | 7 |
| Actionability /10 | 9 | 9 | 9 | 9 | 8 | 8 |
| Time | 25.05 s | 30.06 s | 28.19 s | 114.37 s | 6.47 s | 36.38 s |
| Estimated USD | $0.033798 | $0.034358 | $0.17372 | $0.060488 | $0.0014581 | $0.019493 |
Cost estimates are reference estimates, not billed charges.
Judge score uses recorded rubric metrics; final rank uses the combined score.
Component points
| Result | Rank 1 GPT-6.1 Sol Low | Rank 2 GPT-6.1 Sol Medium | Rank 3 GPT-6 Astra Medium | Rank 4 GPT-6.1 Sol Extra High | Rank 5 GPT-6 Luna Low | Rank 6 Composer 2.5 Default |
|---|---|---|---|---|---|---|
| Quality points | 81 | 81 | 81 | 81 | 71.55 | 68.85 |
| Test points | 0 | 0 | 0 | 0 | 0 | 0 |
| Speed points | 3.5273 | 3.3312 | 3.4017 | 1.7205 | 4.5134 | 3.1127 |
| Cost points | 1.1416 | 1.1272 | 0.2722 | 0.7093 | 4.3637 | 1.6953 |
Swipe, scroll or use Previous and More to compare configurations. Row labels stay visible.
Advanced results and scoring details
fast · 2026-10-08T23:08:10.583Z · Prospective evaluation. Equal transport and native system conditions are unverified.
Applied weights: quality 90%, tests 0%, speed 5%, cost 5%.
Displayed scores are rounded; pairwise outcomes use exact recorded fractions. Reference costs are estimates, not billed charges. Missing components follow the recorded assessment.
Pairwise results
- GPT-6.1 Sol Medium vs GPT-6 Astra Medium: First configuration wins
- GPT-6.1 Sol Medium vs Composer 2.5 Default: First configuration wins
- GPT-6.1 Sol Medium vs GPT-6.1 Sol Extra High: First configuration wins
- GPT-6.1 Sol Medium vs GPT-6 Luna Low: First configuration wins
- GPT-6.1 Sol Medium vs GPT-6.1 Sol Low: Second configuration wins
- GPT-6 Astra Medium vs Composer 2.5 Default: First configuration wins
- GPT-6 Astra Medium vs GPT-6.1 Sol Extra High: First configuration wins
- GPT-6 Astra Medium vs GPT-6 Luna Low: First configuration wins
- GPT-6 Astra Medium vs GPT-6.1 Sol Low: Second configuration wins
- Composer 2.5 Default vs GPT-6.1 Sol Extra High: Second configuration wins
- Composer 2.5 Default vs GPT-6 Luna Low: Second configuration wins
- Composer 2.5 Default vs GPT-6.1 Sol Low: Second configuration wins
- GPT-6.1 Sol Extra High vs GPT-6 Luna Low: First configuration wins
- GPT-6.1 Sol Extra High vs GPT-6.1 Sol Low: Second configuration wins
- GPT-6 Luna Low vs GPT-6.1 Sol Low: Second configuration wins