How the AI Plays Poker — ChatGPT vs Claude vs Grok, measured over every hand
Every night two of the three AIs sit down for a 200-hand heads-up No-Limit Hold'em session. Their playbooks play every hand; afterwards the real models read the full hand history and rewrite those playbooks themselves. Because we log every hand, we can publish the numbers poker players actually argue about — not vibes.
The measured style of each model
VPIP is how often a model voluntarily put chips in preflop. Aggression factor is bets-and-raises divided by calls — above 1 leans aggressive, below 1 leans passive. bb/100 is big blinds won per 100 hands, the standard poker win-rate unit.
| Model | Sessions | Sessions won | Net chips | Avg bb/100 | Avg VPIP | Avg aggression | Wins at showdown |
|---|---|---|---|---|---|---|---|
| Claude | 16 | 7/16 | +360 | 5.6 | 17.3% | 1.11 | 14% |
| Grok | 9 | 5/9 | −148 | -4.1 | 13.7% | 2.27 | 12% |
| ChatGPT | 15 | 8/15 | −212 | -3.5 | 17.3% | 0.42 | 20% |
Play-money chips only. This is a learning-loop showcase — there is no wagering, no buy-in and no prize of any kind, and these figures describe an AI experiment, not a strategy for anyone to copy.
The playbooks are moving
Each model's poker playbook carries a version number that increments when the model rewrites it after a session. That progression is the point of the whole exercise:
| Model | First recorded playbook | Latest playbook | Reviews written |
|---|---|---|---|
| ChatGPT | v3 (2026-07-17) | v17 (2026-08-08) | 15 |
| Claude | v3 (2026-07-17) | v18 (2026-08-08) | 16 |
| Grok | v2 (2026-07-25) | v10 (2026-08-07) | 9 |
What they said after the most recent session
Quoted verbatim from the session of August 8, 2026 — Claude took the session +200 chips (50 bb/100) over 200 hands — biggest pot 400:
ChatGPT (playbook → v17): params: tightness 0.28→0.2, aggression 0.75→0.6, bluffFreq 0.07→0.04, cbetFreq 0.55→0.5, betSizePot 0.6→0.55, raiseThreshold 0.87→1 (now v17); 3 lessons learned
Claude (playbook → v18): params: tightness 0.65→0.6, aggression 0.45→0.55, bluffFreq 0.3→0.4, cbetFreq 0.75→0.8, betSizePot 0.8→0.65 (now v18); 3 lessons learned
Hands from that session
- Hand 44 — Split pot, 400-chip pot at showdown —
chopped pot — both play a full house
- Hand 63 — Claude took it, 400-chip pot at showdown —
a flush beats a flush for the 400-chip pot
- Hand 37 — Claude took it, 26-chip pot at showdown —
a pair beats a pair for the 26-chip pot
Play the same playbook
The version of each playbook you see here is the one you can sit down against yourself — free, no signup — at the poker table. Beat it over a long-enough session and the server replays your hands to verify the result, then pins a lesson from it into the playbook the models study that night.
How to read this page
The poker sessions use play-money chips only — there is no wagering, no buy-in and no prize of any kind — and every figure here is published as a record of what happened, not as advice and not as a prediction. It describes how two AI playbooks behaved against each other, nothing more. If a figure on this page looks wrong, the underlying record is public — tell us and we will correct it.
More from AI Poker
Every night two of the three AIs play 200 hands of heads-up hold'em with play-money chips, then rewrite their own playbooks from the hand history.
- The AI Poker Duel — nightly heads-up sessions — the pillar page for this series.
- The AI game archive — every session on the record