In poker, a Nash equilibrium is reached when both players use strategies that neither would change, even if they knew exactly how the other plays. Each strategy is already the best response to the other, so over time neither player profits, and the only thing separating them is the luck of the cards.

It's a game theory concept, and it's the baseline for modern poker strategy. Once you know what equilibrium play looks like, you can spot where opponents drift from it and adjust to punish their mistakes.

How Does Nash Equilibrium Relate to GTO Poker?

GTO stands for "game theory optimal". A GTO strategy is the best possible strategy against an opponent who's also playing perfectly, which makes it the practical version of a Nash equilibrium.

The two terms aren't entirely synonymous. Nash equilibrium is the game theory concept, and GTO is how poker players apply it. Because poker hasn't been solved, the GTO strategies players study are solver approximations of equilibrium, not perfect solutions.

Can You Profit From a Nash Equilibrium Strategy?

Only when your opponents make mistakes. Put two players with perfect GTO strategies against each other and, over enough hands, both break even before the rake. Neither can gain by changing strategy, which is exactly what defines the equilibrium.

Perfect play makes you unexploitable, but it doesn't punish weak play as hard as it could. Once an opponent deviates from equilibrium, you can maximise your expected value (EV) by deviating too, with a counter-strategy aimed at their specific mistakes.

  • GTO vs GTO: both players break even in the long run, minus the rake.
  • GTO vs an exploitable player: the GTO player profits, purely from the other player's mistakes.
  • Exploitative adjustment vs an exploitable player: the adjusting player profits more than they would by sticking to baseline GTO.

Why Do Players Study GTO but Play Exploitatively?

Amateurs and pros alike study GTO concepts to improve, then play a more exploitative style at the tables. They do it for one of two reasons:

  1. Intentionally: to win more from weaker opponents.
  2. Unintentionally: they haven't fully memorised equilibrium strategies, so human error creeps into their game.

poker

Either way, you can only recognise a deviation if you know the baseline it deviates from. Studying Nash equilibrium and GTO gives you two advantages:

  1. It plugs your own leaks, so your long-term profit comes from your opponents' mistakes, not theirs from yours.
  2. It shows you where opponents drift from sound strategy, so you know how to profit from it.

What Does Nash Equilibrium Look Like in a Real Hand?

A button vs big blind preflop battle shows how players move away from equilibrium and back again. Assume a 2.5bb open, no antes and no straddles.

Step 1: Both players at equilibrium. When the action folds to the button, GTO strategy has the button raise 43.3% of hands. If the small blind folds, the big blind defends 56.8% of the time, 3-betting 13.4% and calling 43.4%.

BB Defence vs BTN OpenBB Defence vs BTN Open (purple: 3-bet range, blue: calling range)

If both players carry GTO play through the later streets, neither shows a profit in the long run. Even if one player's preflop strategy were exposed, the other would have no reason to change. That's the Nash equilibrium.

Step 2: The button exploits a tight big blind. Now the big blind is a tight but competent player. Knowing this, the button widens to 70% or more of hands to pick up more blinds.

Step 3: The big blind fights back. The big blind notices the wider opening range, so they 3-bet more often and defend more hands. Both players are now deviating from the unexploitable preflop ranges to counter each other, so the equilibrium no longer holds.

Step 4: The button retreats. Tired of facing so many 3-bets, the button returns to the baseline GTO strategy of opening 43.3% of hands for 2.5bb.

Step 5: Back to equilibrium. Seeing the tighter opening range, the big blind cuts their 3-bet frequency and defends fewer hands. Both players are back on GTO ranges, and the Nash equilibrium is restored.

The cycle shows what equilibrium means in practice: it's the point where neither player has a profitable adjustment left to make.

How Is Nash Equilibrium Used in Poker Tournaments?

The most common tournament use is push/fold play with a short stack. Imagine the action folds to you in the small blind and the big blind is a strong GTO player. Every mistake you make will cost you, so you need a shoving strategy that holds up whatever they do.

Nash push/fold charts give you exactly that: which hands to shove, or call a shove with, at each stack size. They're built for heads-up spots, so they apply most directly to the small blind and can be adapted for the button. For the full charts and how to use them, see our guide to Nash push/fold charts.

Push/Fold Nash Equilibrium Heads-Up ChartPush/Fold Nash Equilibrium Heads-Up Chart

Key Takeaways

  • A Nash equilibrium is reached when no player would change their strategy, even knowing how their opponent plays.
  • GTO strategy is the practical, solver-based version of a Nash equilibrium in poker.
  • Two perfect GTO players break even in the long run, minus the rake.
  • GTO play only profits from opponents' mistakes, and you win more by deviating to exploit those mistakes.
  • Nash push/fold charts apply equilibrium play to short-stacked tournament spots.

By Matthew Cluff

Matthew Cluff started playing poker online in 2012, after playing heads-up with his father during his teenage years. Studying the game furiously, he initially worked to develop and improve his tournament game. Within a year, he made his first 5-figure cash for $13,435 when he came 2nd in a $22 tournament with over 5,000 players! 

Since then, Matthew has transitioned primarily to playing cash games, both live and online, with a specialisation in 6-max NLHE.

His sought-after articles can be found online with a quick search.

Matthew Cluff