MATED
← Blog

Deliberate practice in chess: what transfers, what doesn't

A working definition of effective chess practice, the parts of the deliberate-practice idea that don't survive contact with the board, and the error data that tells you where to point.

·8 min read·Includes figures from Mated’s own move data

The two-sentence answer

Practice the specific thing that is losing you games, at a difficulty where you fail regularly, with the correct answer arriving within seconds of your attempt. Volume, schedule and choice of app are all downstream of those three conditions, and none of them fix a session that fails them.

The rest of this is about which parts of the deliberate-practice idea actually apply to chess, which parts get quoted at you and shouldn't be, and what your own error distribution says about where to aim.

What deliberate practice requires, minus the mythology

The concept comes from Ericsson's work on musicians and later on chess players. The Charness et al. study of chess expertise (2005) found that serious solitary study was the strongest predictor of rating among the activities they measured — stronger than tournament play. That is correlational and the hours were self-reported, so treat it as suggestive rather than proven.

The mechanism, stripped of the folklore, is four conditions:

  • The task sits just past what you can currently do. Not comfortably inside it, not miles beyond it.
  • Feedback is fast and specific. "Wrong" is useless. "Wrong, because ...Bb4+ wins the exchange" is practice.
  • You repeat the corrected version. One-off exposure to a mistake is not learning.
  • Attention is undivided. A puzzle rush with the football on is entertainment.
  • The 10,000-hour figure is not part of this. It was a rounded average from one study of violinists and it was never a threshold.

Your errors are not where you think, and they are not evenly spread

Across 219,622 engine-classified moves from 6,954 games analysed on Mated, 3.8% of moves played are outright blunders and another 4.2% are mistakes. That is roughly one move in thirteen being a real error, before you count the 18.4% classed as inaccuracies. In a 40-move game you are making three moves that change the assessment of the position, and you probably notice one of them.

The same set shows the errors clustering hard. 9.1% of middlegame moves are a mistake or worse, against 2.4% in the endgame. That does not mean your endgames are good — you reach fewer of them, and the ones you reach are often already decided. It does mean that if you are choosing between an hour of rook endgame technique and an hour on middlegame candidate moves, the middlegame is where the moves you actually play are going wrong.

There is one more number worth sitting with. In that set the average move gives away 94 centipawns against the engine's best, but the median is far lower, because a small number of very bad moves carry most of the damage. Your average is not being dragged down by a steady drip of slightly-second-best moves. It is being dragged down by three catastrophes. Practice that shaves 15 centipawns off your typical move is aimed at the wrong tail.

If you solve almost every puzzle, they are the wrong puzzles

Across 2,044 rated puzzle attempts on Mated, 94.1% were solved first time. That is a pleasant number and a bad one. A set pitched at the edge of your ability should be failing you regularly; a 94% success rate means most of the attempts were confirming a pattern you already owned.

What the right failure rate is, nobody has established for chess. There is work in machine learning and simple perceptual tasks pointing at roughly 85% accuracy as the optimal training point, and it is genuinely unclear whether that transfers to a task as compositional as calculating a variation. My own guess, and it is a guess, is that if you are solving four out of five without a struggle, raise the rating until you aren't.

The practical fix is boring: stop letting the puzzle server pick easy tactics from your rating band and start solving positions where you already went wrong. A position from your own game at move 22, with the eval swing hidden, is harder than a themed pin puzzle because you have already demonstrated you can't solve it. That is what Mated builds its daily set from — the positions your engine pass flagged, not a generic queue.

The parts that don't transfer

Deliberate practice was described for skills with a clean feedback loop and a coach standing next to you. Chess breaks several of those assumptions and it is worth being honest about which.

No coach. Most of the literature assumes someone who can see the flaw you can't. An engine is not that: it tells you the move was bad, not why you chose it. The gap between "Qxd5 loses to Bb4+" and "I stop calculating as soon as I see material" is the entire coaching job, and you have to do it yourself.

No clean repetition. A violinist plays the bar again. You cannot play the position again without knowing the answer. The nearest substitute is finding other positions with the same error class, which is more work than it sounds.

The "practice must be unpleasant" line. Ericsson's point was that deliberate practice isn't inherently rewarding, not that misery is evidence of progress. Plenty of strong players enjoy analysis. Do not use tedium as a quality signal.

Opening study is the biggest false transfer. It satisfies every surface feature of deliberate practice — repetition, clear right answers, measurable recall — and below about 1800 it rarely touches the moves that are costing you, because your losses are starting after the theory runs out. Learn enough to reach a playable middlegame and stop.

One worked example instead of a checklist

Take a game you lost this week. Find the single largest eval swing and stop there. Say it was move 23: you played Rxd7, having calculated Rxd7 Qxd7 Bc6, skewering the queen and rook. You played it in eight seconds. He answered Qe1+ and mated on the back rank.

The useless conclusion is "I blundered." The useful one is a sentence about your process: I stop calculating when I find a line that wins material, and I do not check his forcing replies. That is an error class, and error classes are what you can drill.

The drill writes itself. Pull twenty positions out of your last thirty games where you had a capture available. For each one, do not find the best move. Answer one question only: what is his most forcing reply? Thirty seconds each, ten minutes total. You are not practising tactics, you are practising the habit that failed — looking at his move before committing to yours.

That drill takes an afternoon to build by hand. This is the one place a tool earns its keep: running an engine over your archive and grouping the flagged positions by what went wrong is exactly what Mated automates, and it is why the daily session is fifteen minutes rather than an hour of setup.

What to measure, given that rating won't tell you

Rating moves too slowly and too noisily to tell you whether last month's practice worked. Fifty games at 1400 will swing 80 points on variance alone. If you change your training and your rating goes up, you have learned almost nothing about the change.

Track the error rate instead, split by phase. Blunders per game, mistakes per game, and where in the game they land. Those numbers move before rating does and they respond to a specific intervention in a way rating doesn't. If you drilled forcing replies for three weeks and your middlegame blunder rate hasn't moved, the drill is wrong, or you aren't doing it, and you know that in three weeks rather than three months.

The honest limit: the numbers in this article describe where errors occur across a large set of games. They do not show that targeted practice reduces those errors faster than just playing more. Nobody has run that comparison properly, on any platform, as far as I know. What the data supports is where to aim. Whether aiming helps as much as the coaching world assumes is still an open question, and you should be suspicious of anyone who tells you otherwise with a percentage attached.

Aim your practice at the small number of very bad moves that cause most of the damage, not at your average move, and check your failure rate rather than your rating to see whether it's working.

Questions

How long should a practice session be?
Long enough to hold full attention, which for most adults after a working day is 15 to 30 minutes. A 90-minute session where the last hour is done half-asleep is worse than 20 focused minutes, because you are rehearsing sloppy calculation. If you have more time, play a long game rather than extending the drilling.
Is blitz ruining my chess?
No, but it isn't practice either. Blitz exercises pattern recall and clock handling, and it does not exercise the thing that produces most of your damage — checking the opponent's replies before you commit. It's fine as play. Just don't count it towards your training hours and then wonder why nothing changed.
How many puzzles a day?
Fewer than you're doing, harder than you're doing. Ten puzzles where you fail three and understand all ten afterwards beats fifty where you fail two and review none. If your first-time solve rate is above about 85%, the set is too easy to be teaching you anything.
Should I analyse my games with the engine on or off?
Off first, then on. Go through the game yourself and write down where you think it went wrong, then check. The gap between your guess and the engine's answer is the actual information — it tells you what you're blind to. Turning the engine on immediately gives you a list of moves and no idea why you didn't see them.
I'm 1200 and I keep losing on time in won positions. Is that a practice problem?
Partly. Time trouble is usually a decision-procedure problem: you spend four minutes on moves with one reasonable candidate and eight seconds on the critical one. Practising faster calculation rarely fixes it. Practising a triage habit — is this position forcing, and if not, play the obvious move in 15 seconds — usually does.

See this in your own games

Mated reads your last games, scores them across five categories and builds a fifteen-minute session out of what it finds. No card.