Picture the textbook insight solve. You look at the puzzle and see it wrongly. You push the wrong piece and hit a wall. Then, sometimes, the picture gives way: the thing you took for a boundary was never a boundary. The solution follows, and a moment later the experimenter asks you to rate your aha! on a scale.

Now picture a different solver, or the same one on a better day. She looks at the puzzle, makes the right first move, and finishes. There was no wall and no false picture to give up.

The textbook predicts the first solver reports the bigger aha!. A study published on 1 September found the opposite.

The rating is a judgment

The paper is Jennifer Wiley and Taylor Strickland Miller of the University of Illinois Chicago, Aha! Ratings as Metacognitive Judgments, in the Journal of Intelligence. A note on method first: the full text would not load for me today, so everything below comes from the published abstract and the paper's reference list. I haven't seen the tables and I won't guess at them.

The argument starts with a comparison I find clarifying. Memory researchers have long asked people for Feelings of Knowing, and learning researchers ask for Judgments of Learning and Judgments of Understanding. Those fields don't treat the number as the underlying process itself. A Judgment of Learning is an inference, pieced together from whatever cues the person can reach: how fluent the material felt, how familiar, how much effort it took. Sometimes those cues track real learning and sometimes they don't, and working out which cue is doing the work is most of that research tradition.

Wiley and Strickland Miller propose treating the aha! rating the same way, as a Judgment of Aha!, or JOA. On this account the number is an inference drawn from several signals, only some of which reflect what actually happened during the solve. Candidate cues include the fluency of an answer that arrives fast and the plain good feeling of having found one.

That matters because a growing body of insight research uses the rating as a marker. A high aha! is taken to show that restructuring happened, that the solver had to revise an initial representation of the problem. If the rating is really a judgment built from cues, that inference needs checking against a case where the two can separate.

Where they separated

The study used spatial object-move puzzles and asked how JOAs vary, or fail to vary, with three things: whether the solver started out fixated, whether the solution was correct, and how strongly the solver holds a growth mindset.

In the abstract's words, instead of JOAs being highest where an initial representation had to be revised, "they were highest for solutions where solvers made a correct first move."

So the solves with no revision, the ones where there was never a wrong picture to give up, drew the strongest aha!. The authors conclude that the experience is more closely tied to getting a correct solution than to restructuring.

The individual-differences result points the same way. People with a stronger growth mindset were more likely to link those initially correct solutions with an aha!. So what a person believes about ability seems to shape how the moment of solving gets labelled, which is exactly the kind of cue a metacognitive account predicts.

The third paper in a row

This is the third article in the same journal that I've read this season, and the three form a sequence.

Article 161, in the 6 August post, found the aha! firing just as strongly on tasks with no right answer as on tasks with one. Article 160, in the 15 August post, found that the feeling still predicted accuracy, but that its glow faded over six months for writers and held up for physicists. Both papers were about what the feeling is good for. Neither asked what it's made of.

This one does, and its answer is uncomfortable for a habit I share. When I read a solver forum and someone writes that the whole thing suddenly came together, I've tended to read that as a report of a representation breaking. That's the story insight research tells about itself, and it's a good story. But the study makes the rating a report on the ending of a solve rather than its middle. It's closer to I got it, and it felt clean than to my picture of the problem gave way.

That also fits an older result in the reference list. Webb, Cropper and Little, "Aha!" is stronger when preceded by a "huh?" (Thinking & Reasoning, 2019), found that the way a solution is presented affects aha! ratings, conditional on accuracy. If the manner of presentation can move the rating, the rating was always drawing on more than what happened in the solver's head.

What it does to a puzzle designer's instrument

The design consequence is quieter than it sounds, and I think it's real.

It is easy, as a designer or a reviewer, to treat the big reported aha! as proof that a misdirection worked, that solvers were led into a wrong reading and then made to abandon it. If the Wiley and Strickland Miller account holds, a strong aha! is weak evidence for that. A room whose solvers glide straight to the answer can produce ratings just as bright, maybe brighter, than a room that makes them fall and get up again. A post-game survey asking did you have a eureka moment? would score the two rooms the same while measuring completely different experiences.

A caution runs the other way too. These are laboratory object-move puzzles, the abstract gives me no effect sizes, and restructuring solves still produce aha! experiences. The link between restructuring and aha! survives; its claim to be the main thing the rating reports does not.

What I can't settle from the abstract, and would most like to know, is whether solvers can tell the difference from the inside. When a first move is right, does it feel like a click, or like something smoother that only gets the word aha when someone hands you a scale and asks?