Skip to content
QuirkLab

The Dunning-Kruger Effect: What It Actually Says

The curve everyone shares does not appear in the paper it claims to depict.

7 min readParadoxesChecked against primary sources
A short ladder with only three rungs standing against a pale wall.

The famous chart showing a confidence spike among the incompetent — Mount Stupid — appears nowhere in Kruger and Dunning's 1999 paper. It was invented later. And the effect itself is mostly a statistical artefact: the pattern can be reproduced from random noise plus regression to the mean. Something real does survive, but it is not what the meme says.

This may be the most-cited psychology finding of the last thirty years, and almost every popular version of it is wrong in at least two ways.

What the 1999 paper actually said

Justin Kruger and David Dunning published "Unskilled and unaware of it" in the Journal of Personality and Social Psychology. Participants took tests of humour, grammar and logical reasoning and then estimated their own performance.

The reported pattern: people scoring in the bottom quartile substantially overestimated their rank, while top performers slightly underestimated theirs. The proposed explanation was metacognitive — the skills needed to perform well are the same skills needed to recognise you are performing badly, so the incompetent lack the tools to see it.

Note what is not there. No confidence peak among beginners, no valley of despair, no curve rising toward a plateau of expertise. The paper contains a fairly simple pair of lines.

Mount Stupid is an invention

The chart everyone shares — confidence spiking to a peak labelled something like Mount Stupid, crashing into a Valley of Despair, then climbing a Slope of Enlightenment — does not appear in Kruger and Dunning's work or in Dunning's follow-up research.

It was retrofitted onto the finding, seemingly through management and self-improvement material in the mid-2000s, and then spread far more widely than the research it claims to depict.

So the single most recognisable image associated with this effect is not a finding at all. That alone puts it in the same category as other claims that circulate on repetition rather than evidence — the mechanism behind something appearing everywhere once you have noticed it.

The statistical problem

The deeper issue is with the analysis, and it took nearly two decades to be widely understood.

Edward Nuhfer and colleagues published critiques in 2016 and 2017 arguing that the characteristic pattern can be reproduced using random data. Gilles Gignac and Marcin Zajenkowski reached a similar conclusion in 2020 in a paper titled, plainly, "The Dunning-Kruger effect is (mostly) a statistical artefact."

Two mechanisms do most of the work. The first is regression to the mean: any measurement contains noise, so people with extreme scores will on average have self-assessments closer to the middle — not because of insight but because of arithmetic. The second is the better-than-average effect, the well-documented tendency for most people to rate themselves above average regardless of skill.

Combine a general tendency to self-rate near or above the middle with regression to the mean, and you generate the Dunning-Kruger pattern from data containing no metacognitive deficit whatsoever.

The autocorrelation problem

There is a third and subtler issue. The original analysis sorts participants by test score and then plots self-assessment against that same score, which builds a relationship into the chart before any psychology is involved.

Gignac and Zajenkowski showed that when continuous-variable methods replace the original quartile-based approach, the effect shrinks substantially. The size of the reported effect depended heavily on how the data were carved up.

One more detail is worth knowing about the original design. Participants estimated their percentile rank — a comparative judgement about where they sat relative to everyone else. Most people have very little information about the distribution they are being compared against, so the task asks for a number they have no basis to produce, and the middle is the natural default.

What survives

This is where a responsible correction has to be careful, because the strong debunking is also an overreach.

People genuinely are poor at assessing their own ability. The classic meta-analytic estimate from Mabe and West found a correlation of around 0.29 between self-assessed and objectively measured performance — a real relationship, and a weak one. Self-assessment carries some signal and a great deal of noise.

Even critics of the effect concede that self-assessment accuracy correlates with actual skill. The disagreement is about magnitude and about whether a specific metacognitive deficit at the bottom of the distribution is required to explain the data. The parsimonious statistical account fits without it.

So the defensible claim is narrower and duller than the meme: self-knowledge about ability is unreliable for almost everyone, and the dramatic incompetence-breeds-confidence version is largely an artefact of how the numbers were handled.

How the two claims differ in practice

It is worth separating what the strong and weak versions each predict, because only one of them is testable in ordinary life.

The strong version says the least skilled are dramatically and distinctively overconfident — a specific deficit located at the bottom of the distribution. The weak version says self-assessment is noisy for everyone, with people at both extremes drifting toward the middle.

These make different predictions about the top of the range. On the strong account, experts are reasonably calibrated and merely modest. On the statistical account, high performers should underestimate themselves by roughly the amount low performers overestimate — which is what regression to the mean produces, and broadly what the data show.

There is also a supply problem in the literature itself. The three main critique papers have collectively attracted a small fraction of the citations of the 1999 original, so even within research the corrected account has not displaced the first one.

Why the wrong version is so appealing

The popular Dunning-Kruger effect is almost always deployed about other people. It explains the overconfident colleague, the commenter, the politician. It is a rare psychological finding that arrives pre-loaded with a target.

That is worth noticing, because it is precisely the shape of a claim that should attract scrutiny rather than agreement. A finding that flatters the person citing it will spread faster than the correction to it — and in this case, the correction has been available in the literature for a decade while the collective citation count of the critique papers remains a small fraction of the original's.

The genuinely useful takeaway is not about the incompetent. It is that your own self-assessment, whatever your skill level, is a weak instrument — which pairs with the finding that a small blunder only helps someone already perceived as competent. How ability is perceived, including by its owner, is a much looser business than either the chart or the intuition suggests.

Same trick, different system
Why You Wake Up One Minute Before Your Alarm

Frequently asked questions

Mostly a statistical artefact. Critiques from 2016 onward showed the characteristic pattern can be reproduced from random data through regression to the mean and the better-than-average effect, without any metacognitive deficit being involved.

This article is educational science trivia about everyday human biology and psychology. It is not medical advice, diagnosis, or treatment, and it is not a substitute for care from a qualified professional.

More from paradoxes