Trap/Cognitive Bias/No. 0489

Illusory Superiority

Illusory superiority is the tendency for people to rate themselves above a comparison group on desirable traits or skills. Studied in self-evaluation research, including Ola Svenson’s work on drivers, it is also called the better-than-average effect and varies with the trait and comparison group.

Also called Better-Than-Average Effect · Above-Average Effect · Lake Wobegon Effect

a trap: easy to walk into

01You've seen this when…

  1. in life

    You read a quiz about listening skills and immediately picture friends who interrupt. You place yourself near the top before considering how often you check your phone while someone talks.

  2. at work

    Several managers rate their communication skills above the department average. One counts quick replies, another counts clear instructions, and another counts keeping everyone informed.

  3. out in the world

    At a neighborhood road-safety meeting, drivers describe themselves as unusually careful. The proposed restrictions sound necessary for the other people on the road.

02The idea

Most people can quickly name something they do better than their peers. Sometimes the evidence backs them up. Illusory superiority appears when favorable self-comparisons systematically outrun what the comparison can support.

The clearest demonstration asks people to rank themselves relative to a group’s median, its midpoint. At most half can fall strictly above it. When a large majority places itself in the upper half, those rankings cannot all be correct.

The effect depends on the question. Being a good colleague might mean meeting deadlines, helping newcomers, speaking candidly or avoiding conflict. Each person can select the standard that makes their own strengths count most. The comparison group matters too: all adults, coworkers and experienced specialists set very different bars.

This is one form of self-enhancement. An above-average rating is accurate for many people; evidence of the bias comes from systematic patterns of inflated relative standing. A single confident answer provides little evidence by itself.

03Why it happens

Several processes can produce the same favorable ranking. Their importance changes with the task and how the question is asked.

  • People choose a flattering yardstick. Broad traits leave room for personal definitions. Someone who gives precise instructions treats precision as the heart of good communication. Someone who listens patiently treats listening as central. Research on ambiguous traits shows how these individual definitions can support favorable self-assessments.
  • Desirable qualities carry emotional weight. Rating oneself as honest, capable or considerate helps preserve a positive self-image. Motivated reasoning can make supporting memories easier to accept and awkward counterexamples easier to explain away. Studies find that desirability and perceived control over a trait affect self-evaluations.
  • Self-knowledge dominates the comparison. People usually have more information about their own performance than about everyone else’s. On an easy task, doing well can feel like evidence of superior ability even when peers also do well. Research also finds below-average judgments on difficult tasks, where personal struggle dominates the estimate.
  • The comparison person stays blurry. An unspecified average person gives the mind considerable room to maneuver. A reader can picture careless drivers, unhelpful coworkers or inattentive friends while recalling their own strongest moments. Naming a specific comparison group makes the claim easier to examine.

04A worked example

In a 1981 study, psychologist Ola Svenson asked participants in the United States and Sweden to judge their driving skill relative to their comparison group. Among the American respondents, 93% placed themselves in the more skillful half. Among the Swedish respondents, 69% did so.

What it looks like Both groups contain an unusually large share of skilled drivers, especially the American group.

What’s actually going on The rankings concern standing within the comparison group. Having 93% of respondents claim a place in its upper half produces incompatible judgments. The study demonstrates inflated comparative self-ratings at the group level. It did not include an objective driving test, so it cannot identify which individuals overestimated their skill or show that their confidence caused crashes.

What would have helped A shared definition of driving skill and comparable performance evidence would make the rankings testable. Ratings could then be checked against the same criteria for everyone. That is a proposed safeguard; Svenson’s study did not test it as an intervention.

05How to spot it

These are prompts for checking the assessment. None proves that a particular person’s high rating is mistaken.

06What to do instead

  • Define the skill before rating it. Replace a broad label such as good communicator with observable criteria: whether recipients understand the request, whether decisions reach affected people, or whether instructions need repeated clarification. Agree on the criteria before anyone sees their score.
  • Name the comparison group. Choose people doing similar work under similar conditions. Comparing a beginner with all adults answers a different question from comparing that beginner with trained specialists.
  • Collect evidence about both sides. Review the same kinds of outcomes for yourself and the comparison group. For a work skill, this might mean comparable assignments or feedback gathered with the same questions. Calibration requires a reference that can reveal error.
  • Look deliberately for contrary cases. Use the consider-the-opposite strategy: identify a recent occasion when a peer handled the same challenge better. Record what they did and which criterion it satisfied.

For a skill you use repeatedly, write down a prediction before the next attempt and compare it with the result. Over time, this creates evidence that a flattering recollection cannot easily rewrite.

07When it isn’t illusory superiority

Expertise can justify high relative confidence. A trained interpreter may accurately rate their language skills above those of most adults. The relevant questions are whether the group fits the claim and whether the evidence supports the ranking.

The word average also needs care. A majority can exceed an arithmetic mean when a few very low scores pull that mean downward. The mathematical contradiction applies most cleanly to claims about the upper half or above the median.

Overconfidence covers a broader range of errors, including exaggerated absolute performance and excessive certainty. The Dunning-Kruger effect concerns how performance level relates to errors in self-assessment. Illusory superiority can occur across ability levels.

Some qualities, including kindness and integrity, have no single agreed score. Clearer definitions can expose disagreements while leaving legitimate differences in values.

08Roots

Ola Svenson chose driving as a setting for studying judgments about skill and risk. Drivers make consequential decisions every day, yet most receive little standardized feedback about their standing among peers. A trip completed without incident gives limited information about comparative ability.

His 1981 study became a memorable demonstration within a longer tradition of self-evaluation research. Asking drivers to locate themselves within a group turned a comfortable personal impression into a ranking that could be checked for consistency. The crowded upper half made the puzzle visible.

Later researchers examined what people meant by the qualities they were rating. Mark Alicke studied how desirable and controllable traits shaped self-evaluation. David Dunning and colleagues investigated how people used their own definitions of ambiguous abilities. Those studies helped explain why changing the wording could change the apparent advantage.

The idea traveled into discussions of management, education and everyday judgment as the better-than-average effect. A comprehensive meta-analysis published in 2020 brought together evidence across studies and examined how the effect varied with traits, comparison methods and other conditions.

09How solid is this?

ContestedMixedUsefulEstablished

A comprehensive meta-analysis supports a reliable tendency toward favorable self-comparison, with substantial variation across traits, tasks and methods. Self-ratings establish patterns of claimed standing; comparable performance evidence is needed to assess an individual’s error.

10Connections

confused withconfused withcountered bycountered byfollows frompart ofpart ofIllusorySuperiorityBias Blind SpotDunning-KrugerEffectConsider-the-OppositeStrategyCalibrationMotivatedReasoningOverconfidenceEffectNot written yetSelf-EnhancementBiasNot written yetFalse UniquenessEffectNot written yetSocialComparison Theory

11Origin and sources

Developed within self-evaluation and social comparison research. Ola Svenson’s 1981 driving study is a classic demonstration; later work by Mark Alicke and David Dunning examined conditions that shape favorable self-comparisons.

  1. [1]Svenson, O. (1981). Are we all less risky and more skillful than our fellow drivers? Acta Psychologica, 47(2), 143–148.
  2. [2]Alicke, M. D. (1985). Global self-evaluation as determined by the desirability and controllability of trait adjectives. Journal of Personality and Social Psychology, 49(6), 1621–1630.
  3. [3]Dunning, D., Meyerowitz, J. A., & Holzberg, A. D. (1989). Ambiguity and self-evaluation: The role of idiosyncratic trait definitions in self-serving assessments of ability. Journal of Personality and Social Psychology, 57(6), 1082–1090.
  4. [4]Moore, D. A., & Small, D. A. (2007). Error and bias in comparative judgment: On being both better and worse than we think we are. Journal of Personality and Social Psychology, 92(6), 972–989.
  5. [5]Zell, E., Strickhouser, J. E., Sedikides, C., & Alicke, M. D. (2020). The better-than-average effect in comparative self-evaluation: A comprehensive review and meta-analysis. Psychological Bulletin, 146(2), 118–149.

Suggest an edit· Updated 2026-10-02