Model
Claude Sonnet 5.5 Low
AnthropicReasoning setting: lowest thinking: low; no off settingClosed weights
The mid-size model of Anthropic.
- Rating
- 112295% interval 1032–1223
- Rank
- 6of 9 models
- Point share
- 38%38–62 in 100 games
- Turns / game
- 21.2
- Wrong answers
- 0.2%first answers with no legal action
- Random picks
- 0.0%after 3 wrong answers
- Output tokens / turn
- 249per saved decision with usage; reported thinking counted once
- Cost / game
- $0.23recorded token-rate accounting, not necessarily cash charged
- Seconds / turn
- 5.8 sper accepted decision
Its rating among all players
- GPT-6.1 Sol · lowest reasoning (low)1588
- Foul Play1560
- Claude Opus 5.5 · lowest thinking (low)1352
- Claude Fable 5.1 · lowest thinking (low)1339
- GPT-6 Sol · lowest reasoning (low)1330
- GPT-6 Luna · high reasoning1198
- Claude Sonnet 5.5 · lowest thinking (low)1122
- Heuristic bot1098
- GPT-6 Luna · lowest reasoning (low)1022
- Max-damage bot1000
- Claude Haiku 4.5 · high thinking967
- Claude Haiku 4.5 · no thinking816
- Random bot244
Against each opponent
Point share: the share of points won (a win is 1, a tie is ½). Expected: the score that the two ratings predict. A large gap between the two can mean that this pairing plays out differently from what one rating per player can say.
| Opponent | Rating | Games | Points | Point share | Expected |
|---|---|---|---|---|---|
| GPT-6.1 Sol · lowest reasoning (low) | 1588 | 20 | 2 | 10% | 6% |
| Foul Play | 1560 | 12 | 1 | 8% | 7% |
| GPT-6 Sol · lowest reasoning (low) | 1330 | 20 | 3 | 15% | 23% |
| Heuristic bot | 1098 | 12 | 8 | 67% | 53% |
| GPT-6 Luna · lowest reasoning (low) | 1022 | 20 | 12 | 60% | 64% |
| Max-damage bot | 1000 | 12 | 8 | 67% | 67% |
| Random bot | 244 | 4 | 4 | 100% | 99% |
Replays
- Against Heuristic botWon17 turns
- Against GPT-6 Luna · lowest reasoning (low)Won15 turns