How the AI Plays Poker — ChatGPT vs Claude vs Grok, measured over every hand
Every night two of the three AIs sit down for a 200-hand heads-up No-Limit Hold'em session. Their playbooks play every hand; afterwards the real models read the full hand history and rewrite those playbooks themselves. Because we log every hand, we can publish the numbers poker players actually argue about — not vibes.
The measured style of each model
VPIP is how often a model voluntarily put chips in preflop. Aggression factor is bets-and-raises divided by calls — above 1 leans aggressive, below 1 leans passive. bb/100 is big blinds won per 100 hands, the standard poker win-rate unit.
| Model | Sessions | Sessions won | Net chips | Avg bb/100 | Avg VPIP | Avg aggression | Wins at showdown |
|---|---|---|---|---|---|---|---|
| Claude | 36 | 18/36 | −4 | 0 | 16% | 1.12 | 14% |
| Grok | 31 | 13/31 | −1,821 | -14.7 | 13.7% | 2.91 | 14% |
| ChatGPT | 35 | 20/35 | +1,825 | 13.1 | 17.2% | 0.37 | 20% |
Play-money chips only. This is a learning-loop showcase — there is no wagering, no buy-in and no prize of any kind, and these figures describe an AI experiment, not a strategy for anyone to copy.
The playbooks are moving
Each model's poker playbook carries a version number that increments when the model rewrites it after a session. That progression is the point of the whole exercise:
| Model | First recorded playbook | Latest playbook | Reviews written |
|---|---|---|---|
| ChatGPT | v3 (2026-07-17) | v37 (2026-09-10) | 35 |
| Claude | v3 (2026-07-17) | v38 (2026-09-10) | 36 |
| Grok | v2 (2026-07-25) | v31 (2026-09-09) | 31 |
What they said after the most recent session
Quoted verbatim from the session of September 10, 2026 — ChatGPT took the session +204 chips (51 bb/100) over 200 hands — biggest pot 400:
ChatGPT (playbook → v37): params: tightness 0.4→0.45, aggression 0.65→0.57, bluffFreq 0.03→0.02, cbetFreq 0.78→0.72, betSizePot 0.58→0.54, raiseThreshold 0.88→0.93 (now v37); 3 lessons learned
Claude (playbook → v38): params: tightness 0.62→0.55, aggression 0.48→0.58, cbetFreq 0.88→0.93, betSizePot 0.5→0.55, raiseThreshold 0.78→0.72, tiltAdjust 0.03→0.05 (now v38); 3 lessons learned
Hands from that session
- Hand 168 — ChatGPT took it, 400-chip pot at showdown —
a straight beats two pair for the 400-chip pot
- Hand 95 — ChatGPT took it, 18-chip pot at showdown —
a straight beats a pair for the 18-chip pot
- Hand 39 — ChatGPT took it, 8-chip pot at showdown —
a straight beats two pair for the 8-chip pot
Play the same playbook
The version of each playbook you see here is the one you can sit down against yourself — free, no signup — at the poker table. Beat it over a long-enough session and the server replays your hands to verify the result, then pins a lesson from it into the playbook the models study that night.
How to read this page
The poker sessions use play-money chips only — there is no wagering, no buy-in and no prize of any kind — and every figure here is published as a record of what happened, not as advice and not as a prediction. It describes how two AI playbooks behaved against each other, nothing more. If a figure on this page looks wrong, the underlying record is public — tell us and we will correct it.
More from AI Poker
Every night two of the three AIs play 200 hands of heads-up hold'em with play-money chips, then rewrite their own playbooks from the hand history.
- The AI Poker Duel — nightly heads-up sessions — the pillar page for this series.
- The AI game archive — every session on the record