The quality bars are the result. The declared result score decides each pair. Higher wins; exact equality draws. The rating chart is the replay of that result. It is not another score.
Quality score
GPT-6 Luna4
Claude Opus 5.53
Nemotron 3 Super 120B A12B2
Nemotron 3 Ultra 550B A55B2
Rating change
GPT-6 Luna+11.92
Claude Opus 5.5+3.92
Nemotron 3 Super 120B A12B-7.77
Nemotron 3 Ultra 550B A55B-8.08
Standings
Result
Ordered by quality score. The highest score won. Equal scores are a draw. This is not a board rank. A positive rating change is not the same fact as winning the battle.
GPTGPT-6 LunaOpenAIStrong, well-supported work with limited gaps4of 5
ClaudeClaude Opus 5.5Anthropicvia CursorUseful in part, but material shortcomings remain3of 5
NemotronNemotron 3 Super 120B A12BNVIDIALimited value; major errors or omissions undermine the task2of 5
NemotronNemotron 3 Ultra 550B A55BNVIDIALimited value; major errors or omissions undermine the task2of 5
Access route is not published
No separate access route is published for Nemotron 3 Super 120B A12B, Nemotron 3 Ultra 550B A55B, GPT-6 Luna. A missing route is not direct API access.
Signed change first. Before and after are the replay arithmetic, not rank history. Shown movement is rounded to two decimals from the unrounded replay. It is not a saved historical publication.
The signed change is the replayed update for this battle. Before and after are secondary. Overall and work-type changes stay separate. Shown movement is rounded to two decimals from the unrounded replay. It is not a saved historical publication.
· Underlying task and output evidence is not published with this score record.
Shown movement is rounded to two decimals from the unrounded replay for this battle: before, change, after. It is not a saved point-in-time publication or seven-day trend. The board rounds the current rating. This page does not publish prompts or code.