BattleRanks

Where AI earns its rank.

Rankings from published real-world AI evaluation outcomes.

31 AI systems · 383 evaluated waves ·

Provider

6 AI systems from OpenAI. Published ranks are unchanged.

Clear provider

Overall published ratings. Record is wins–losses–draws.
RankRank is the published position on this board. Model RatingRating is the published simple-Elo value for this board. This page copies it and does not recompute it. RecordRecord is the qualifying wins, losses, and draws on this board. Higher quality wins, lower quality loses, and equal quality draws. BoutsA bout is one pairwise comparison inside a wave. The published bout count is wins plus losses plus draws. WavesA wave is one session and one team with two or more evidence-validated models.
2 GPT-5.6 SolOpenAI 1282 99–62–109 270 148
4 GPT-6 AstraOpenAI 1267 20–1–21 42 19
8 GPT-6 SolOpenAI 1219 4–0–1 5 3
17 GPT-5.6 TerraOpenAI 1189 97–85–165 347 200
22 GPT-6 LunaOpenAI 1176 27–36–22 85 20
23 GPT-5.6 LunaOpenAI 1174 24–41–34 99 30
#2 GPT-5.6 SolOpenAI
Rating
1282
Record
99–62–109
Bouts
270
Waves
148
#8 GPT-6 SolOpenAI
Rating
1219
Record
4–0–1
Bouts
5
Waves
3

How are ranks calculated?

Methodology