The engine can't teach you. Here is what a question does that +1.3 cannot.
Feedback only works when it lands on a reason. An evaluation is a grade with nothing underneath it. A question makes you commit, and a committed error is the one you remember correcting.
Published 2026-09-03
+1.3 is the most honest thing an engine will ever tell you. It is accurate to a tenth of a pawn, it arrives in milliseconds, and it contains no information about you at all. It does not know why you played the move. It has no theory of what you were trying to do. It cannot tell you which of your habits produced it, and it will not notice when you do it again.
This post is about the difference between a grade and feedback, and about why the specific thing a coach does, asking you a question before telling you anything, is not a courtesy. It is the mechanism.
Three questions a number cannot answer
John Hattie and Helen Timperley's review of feedback research is the one most teachers have read, and its central claim is simple: feedback is among the strongest influences on learning, and its power depends entirely on what it is about. Feedback that tells you how you did on the task, how your process went, and how to regulate yourself next time works. Feedback aimed at the person (“good job”) does almost nothing. Effective feedback, they argue, answers three questions: where am I going, how am I going, and where to next.1
A later meta-analysis found the average effect smaller than Hattie's early syntheses suggested, and hugely variable, for exactly that reason: most feedback is not about the process.2An evaluation answers “how am I going” in the narrowest possible sense and leaves the other two questions blank. It is a score on the outcome of a decision whose reasoning it never saw.
Feedback needs something to land on
Nate Kornell, Matthew Hays and Robert Bjork ran a study in which people had to guess the answer to a question before being shown it, under conditions where the guess was almost guaranteed to be wrong. The failed attempt still improved later memory for the correct answer compared with simply studying it.3Janet Metcalfe's review of the errors literature reaches the same conclusion from many directions: making an error and then receiving the correction beats being handed the answer, and instruction designed to avoid errors is likely counterproductive.4
The mechanism matters for chess. When you commit to a reason, “I took on e3 because trading off that knight looked like the simplest way out,” you have created the thing the correction attaches to. When the coach then shows you that Black recaptures with check and picks up the knight on f7 a move later, the correction has an address. When the engine shows you Qe6 with no commitment on your side, the correction is filed nowhere in particular.
Your surest blunders are the easiest to fix
Brady Butterfield and Janet Metcalfe expected that errors made with high confidence would be the hardest to correct. They found the opposite. On general-knowledge questions, the wrong answers people were most sure of were the ones most likely to be corrected on a later test, once they had seen the right answer.5The finding, called hypercorrection, requires corrective feedback and has mostly been shown with trivia, so we will not stretch it further than that. But it suggests something practical: the move you were certain about, and were wrong about, is the one you will most reliably stop playing, provided somebody makes you state the certainty and then shows you the correction.
That is why a coach asks how sure you were. The engine never will.
The skill the engine gives you no practice at
John Flavell gave the field its name in 1979: metacognition, the knowledge of and monitoring of your own thinking.6A player with good metacognition knows when they are calculating and when they are guessing, notices when a plan has quietly changed, and can tell the difference between “this move is good” and “I want this move to be good.” None of that is a chess skill exactly. All of it decides games.
You cannot practise monitoring your thinking if nothing ever asks you to report it. A question is a monitoring exercise. It forces you to observe what you did and put it into words, which is the only way to find out it was not what you thought.
Even experts benefit from being made to slow down and deliberate. In a study that used engine evaluations to score moves, both strong and weaker players chose objectively better moves after extra deliberation than on their first instinct, on easy and hard problems alike.7
The engine is still there. It is just not the one talking.
To be clear about what we are not saying: passive review is not worthless. In the classic testing-effect study, students who reread a passage did better than self-testers when quizzed five minutes later; the reversal came a week out.8The engine is the most accurate evaluator of a chess position that has ever existed, and Chessalyz runs one underneath everything it does. The point is who speaks first.
eval: +1.3
best: Qe6
Where am I going? How am I going? Where next? Blank, a number, blank.
“You spent a real think on Bxe3. What was your evaluation of the resulting trades, and what other candidate moves were you comparing it against?”
Then, after you have written your reason: where it broke, what Black got for it, and what to look at first next time.
- 1.Hattie, J., & Timperley, H. (2007). The power of feedback. Review of Educational Research, 77(1), 81–112. doi.org/10.3102/003465430298487
- 2.Wisniewski, B., Zierer, K., & Hattie, J. (2020). The power of feedback revisited: A meta-analysis of educational feedback research. Frontiers in Psychology, 10, 3087. doi.org/10.3389/fpsyg.2019.03087
- 3.Kornell, N., Hays, M. J., & Bjork, R. A. (2009). Unsuccessful retrieval attempts enhance subsequent learning. Journal of Experimental Psychology: Learning, Memory, and Cognition, 35(4), 989–998. doi.org/10.1037/a0015729
- 4.Metcalfe, J. (2017). Learning from errors. Annual Review of Psychology, 68, 465–489. doi.org/10.1146/annurev-psych-010416-044022
- 5.Butterfield, B., & Metcalfe, J. (2001). Errors committed with high confidence are hypercorrected. Journal of Experimental Psychology: Learning, Memory, and Cognition, 27(6), 1491–1494. doi.org/10.1037/0278-7393.27.6.1491
- 6.Flavell, J. H. (1979). Metacognition and cognitive monitoring: A new area of cognitive-developmental inquiry. American Psychologist, 34(10), 906–911. doi.org/10.1037/0003-066X.34.10.906
- 7.Moxley, J. H., Ericsson, K. A., Charness, N., & Krampe, R. T. (2012). The role of intuition and deliberative thinking in experts' superior tactical decision-making. Cognition, 124(1), 72–78. doi.org/10.1016/j.cognition.2012.03.005
- 8.Roediger, H. L., III, & Karpicke, J. D. (2006). Test-enhanced learning: Taking memory tests improves long-term retention. Psychological Science, 17(3), 249–255. doi.org/10.1111/j.1467-9280.2006.01693.x
Your last game has a question in it.
Connect Chess.com or Lichess, or paste a PGN. Chessalyz finds the moments that decided the game, asks what you were thinking, and coaches from your answer. Five games a month are free.
Try it on your last game- Why writing down your reasoning makes a chess lesson stick
Fifty years of memory research say the same thing: what you generate yourself, you keep. What you read off a screen, you lose. Here is why a coach asks before they tell.
- Make new mistakes: why your flashcards should be your own blunders
Retrieval beats re-reading, spacing beats cramming, and a position you actually lost beats a puzzle you never met. The research behind flashcards built from your own games.
- Learning by thinking: what a fifteen-minute post-mortem does that another game does not
Masters have always analysed their own games. Cognitive science explains why that habit, not raw volume, is what separates the players who improve from the players who just play.