Model
Claude Haiku 4.5 None
AnthropicReasoning setting: none (thinking off)Closed weights
The small, fast model of Anthropic.
- Rating
- 81695% interval 695–932
- Rank
- 9of 9 models
- Point share
- 14%16–96 in 112 games
- Turns / game
- 21.9
- Wrong answers
- 0.0%first answers with no legal action
- Random picks
- 0.0%after 3 wrong answers
- Output tokens / turn
- 456per saved decision with usage; reported thinking counted once
- Cost / game
- $0.16recorded token-rate accounting, not necessarily cash charged
- Seconds / turn
- 7.3 sper accepted decision
Its rating among all players
- GPT-6.1 Sol · lowest reasoning (low)1588
- Foul Play1560
- Claude Opus 5.5 · lowest thinking (low)1352
- Claude Fable 5.1 · lowest thinking (low)1339
- GPT-6 Sol · lowest reasoning (low)1330
- GPT-6 Luna · high reasoning1198
- Claude Sonnet 5.5 · lowest thinking (low)1122
- Heuristic bot1098
- GPT-6 Luna · lowest reasoning (low)1022
- Max-damage bot1000
- Claude Haiku 4.5 · high thinking967
- Claude Haiku 4.5 · no thinking816
- Random bot244
Against each opponent
Point share: the share of points won (a win is 1, a tie is ½). Expected: the score that the two ratings predict. A large gap between the two can mean that this pairing plays out differently from what one rating per player can say.
| Opponent | Rating | Games | Points | Point share | Expected |
|---|---|---|---|---|---|
| GPT-6.1 Sol · lowest reasoning (low) | 1588 | 20 | 0 | 0% | 1% |
| Foul Play | 1560 | 18 | 0 | 0% | 1% |
| GPT-6 Sol · lowest reasoning (low) | 1330 | 20 | 4 | 20% | 5% |
| Heuristic bot | 1098 | 18 | 3 | 17% | 16% |
| GPT-6 Luna · lowest reasoning (low) | 1022 | 20 | 3 | 15% | 23% |
| Max-damage bot | 1000 | 12 | 2 | 17% | 26% |
| Random bot | 244 | 4 | 4 | 100% | 96% |
Replays
- Against Foul PlayLost14 turns
- Against Random botWon34 turns
- Against GPT-6 Sol · lowest reasoning (low)Won24 turns