Why Your Personality Test Result Feels Like It Was Written About You

By Big Time Trivia Editors. Published , updated .

You finish a personality quiz, open the result, and it lands. It says you can be hard on yourself. It says you sometimes doubt a decision after you've made it. It says you like a certain amount of change and variety, and that you get restless when you're boxed in. You read it twice and recognize yourself in every line.

The feeling is real. Whether it tells you anything about the test is a separate question, and psychologists have been studying it since at least 1949.

One sketch for a whole class

In a paper published in 1949, "The fallacy of personal validation: A classroom demonstration of gullibility", Bertram R. Forer described an exercise with his introductory psychology class. He gave 39 students a personality test called the Diagnostic Interest Blank and told them each would get a brief personality sketch once he had looked over their answers. A week later, each student received a typed sketch with their name on it and rated, on a scale from zero to five, how well it revealed the basic characteristics of their personality.

Every student had received the same sketch, built largely from statements in a newsstand astrology book. Forer's own table of the ratings shows 16 students gave it a 5 and 18 gave it a 4. Four gave it a 3, one gave it a 2, and nobody went lower, which works out to an average of about 4.3. A 1985 review of the research by D. H. Dickson and I. W. Kelly in Psychological Reports, which summarizes Forer's study and the work that followed it, gives the same average.

The sketch had 13 statements, and Wikipedia's entry on the effect reproduces all of them. Three read:

"You have a tendency to be critical of yourself."

"At times you have serious doubts as to whether you have made the right decision or done the right thing."

"You prefer a certain amount of change and variety and become dissatisfied when hemmed in by restrictions and limitations."

Those are the three things this article opened with. Forer also had each student mark every statement as true or false about themselves, or put a question mark if they couldn't tell, and 38 of the 39 marked the first two as true. Thirty-seven marked the third. Forer's name for this kind of check, asking people whether a description fits them, was personal validation, and his warning about it was blunt: "When the inferences are universally valid, as they often are, the confirmation is useless." A description that fits nearly everyone can't separate one person from another, so agreeing with it can't show that the test behind it works.

Why it's named after a circus owner

The name appeared in print seven years later. In a 1956 article in American Psychologist, the psychologist Paul E. Meehl observed that many psychological test reports bore "a disconcerting resemblance" to what his colleague Donald G. Paterson called "personality description after the manner of P. T. Barnum." Meehl proposed, adding that he was "quite serious," that psychologists adopt the phrase "Barnum effect" to stigmatize clinical procedures in which personality descriptions from tests are made to fit the patient "largely or wholly by virtue of their triviality." His reason was practical: "One of the best ways to increase the general sensitivity to such fallacies is to give them a name."

Barnum himself was the circus owner whose formula for success, in Dickson and Kelly's words, was "always to have a little something for everybody." That's exactly what a good Barnum statement does.

What makes a description easy to accept

Dickson and Kelly's review gathers about three decades of follow-up studies. Most were built like Forer's: people take a personality test, wait while it's scored, receive a profile that's supposedly theirs, and rate how accurate it is. In most cases everyone gets the same profile. A few patterns come through.

  • The statements have recognizable shapes. The review quotes a 1955 paper by Sundberg that names four kinds: vague ("you enjoy a certain amount of change and variety in life"), double-headed ("you are generally cheerful and optimistic but get depressed at times"), typical of the group you belong to ("you find that study is not always easy"), and favorable ("you are forceful and well-liked by others").
  • Being told it's yours matters. In a 1972 study by Snyder and Larson, people told that a description had been derived "specifically for them" rated it as more accurate than people told the same description was "generally true of people." Later studies repeated the result, though not every one.
  • Flattery helps. Studies that varied how favorable a description was found the favorable versions were accepted more readily. The review discusses one exception and explains why it doubts that study's method.
  • Credentials matter less than you'd guess. In a 1963 study by Ulrich and colleagues, students gave personality tests to friends and then handed back a generic profile. Comparing the ratings with an earlier experiment in which a psychologist delivered the profile, the researchers concluded the students' version was accepted just as readily. Other comparisons in the review, such as feedback said to come from a computer, a psychologist or a fellow student, found no significant differences. The exception the review names is unflattering feedback, which people accept more readily from someone they see as high in status.

The review also pushes back on the idea that this is simple gullibility. People accept many of these statements, Dickson and Kelly write, "because they do fit, and because they do not have anything else with which to compare them." Many of the statements aren't false. They just aren't about you in particular.

And the effect feeds itself. According to the review, when people find a Barnum profile accurate, their faith in the test that produced it goes up.

Accurate isn't the same as about you

When researchers changed the question, the answers changed too. In a 1977 study by Greene described in the review, people rated a generic description as accurate but did not rate it as a description of their unique personality; asked directly, they could see it would fit their classmates just as well. In a 1984 study by Harris and Greene, people rated Barnum feedback as more accurate than feedback built from their real scores on the California Psychological Inventory, and also as less individual. The generic text beat the real one on accuracy.

A 2008 double-blind study by Alyssa Jayne Wyman and Stuart Vyse makes the same point from another angle. Fifty-two college students were shown two personality summaries, one real and one bogus, and asked to pick their own, once for each of two kinds of test. With summaries drawn from the NEO Five-Factor Inventory, a Big Five questionnaire by Paul Costa and Robert McCrae, they picked their real profile more often than chance would predict. With summaries from computer-generated astrological charts, they couldn't. Yet their accuracy ratings showed a Barnum effect for both.

So a high accuracy rating, on its own, can't tell a real result from a generic one. Picking your description out of a lineup asks a much harder question.

What this means for Myers-Briggs results

When a four-letter type result comes with a paragraph of description, that paragraph is in the same position as Forer's sketch, and the Myers-Briggs framework has its own open questions. Randy Stein and Alexander B. Swan, writing in 2019 about its "immense popularity," concluded that the theory behind the types falters on rigorous theoretical criteria: it lacks agreement with known facts and data, lacks testability, and contains internal contradictions.

That doesn't mean the questions measure nothing. In 1989, Robert R. McCrae and Paul T. Costa found that the four scales of the Myers-Briggs Type Indicator itself measured aspects of four of the five major dimensions of normal personality in the five-factor model, often called the Big Five. What their data didn't support was the idea of distinct types: the scales behaved like four relatively independent dimensions rather than categories you either belong to or don't.

Put those together and the Barnum research gives you a specific caution. A paragraph about your type can feel right for the same reasons Forer's sketch did, whether or not your letters were measured well. To use an invented example, imagine a type description that says you "value authenticity, sometimes feel misunderstood, and have a rich inner life." Hold it up against Sundberg's four kinds and every part is favorable, vague or both: valuing authenticity is a compliment, the "sometimes" makes the second part hard to deny, and a rich inner life is too loose to check. It may be true of you. Nothing in it would stop someone with different letters from accepting it too.

What would show that a test works

Feeling accurate isn't on the list. These are the questions that count, and reading your own result can't answer any of them:

  • Does it give the same answer twice? Take it again in a few weeks. A result that flips on a retake was either close to begin with or isn't picking up something stable.
  • Does it tell people apart? That was Forer's point. A description that fits everyone can't distinguish anyone.
  • Does it predict anything outside the test? This is where trait measures have a real record. A 2007 review by Brent W. Roberts and colleagues considered only prospective longitudinal studies and found that the effects of personality traits on mortality, divorce and occupational attainment were indistinguishable in size from the effects of socioeconomic status and cognitive ability.

How to read your own result

  • Ask the uniqueness question. Not "is this accurate?" but "is this about you in particular, or would it fit most people you know?" Greene's participants, asked whether a generic description captured them as individuals, could see that it didn't.
  • Cross out the Barnum lines. Mark anything vague, anything that has it both ways ("at times you are this, at other times that"), anything that's true of most people around you, and anything that's only a compliment. What's left is the part worth weighing, and it may be short.
  • Try a blind lineup. This is a home version of what Wyman and Vyse did: ask a friend to copy out your type's description and one for a different type, strip the labels, and see whether you can pick yours. It's a rough check, not an experiment, but it asks a better question than whether a description sounds familiar.
  • Ask someone who knows you, about the right things. In a 2010 study, Simine Vazire compared self-ratings with ratings from friends and strangers, and checked all of them against measures drawn from a battery of behavioral tests. People were the best judges of their own neuroticism-related traits, friends were the best judges of intellect-related traits, and self, friends and strangers were equally good at judging extraversion-related traits. Your own view is strongest on traits that are hard to see from outside, like neuroticism; a friend's can be better on traits loaded with judgments of good and bad, like intellect.
  • Read the numbers before the paragraph. A result that shows how your answers split tells you how much the label rests on. A narrow lead on one letter is a different result from a wide one, even when the paragraph underneath is identical.

None of this means the feeling of recognition is fake, or that personality tests are pointless. It means the feeling is where the question starts, not where it ends.

The personality type quiz on this site is built to be read that way. It scores four pairs of opposite preferences, shows the count on each, and flags a close split instead of hiding it behind a letter. Like every quiz here, it hasn't been tested the way research questionnaires are, so give it the same scrutiny this article suggests for any result.

Correction, October 8, 2026: an earlier version of this article dated Forer's study to 1948 and said his sketch was copied nearly word for word from an astrology book. It now follows Forer's 1949 paper, which says the statements came largely from a newsstand astrology book. The earlier version also said research had found that people accept neighboring MBTI type descriptions almost as readily as their own. That claim had no source and has been removed.

Sources

  1. B. R. Forer, The fallacy of personal validation: A classroom demonstration of gullibility, Journal of Abnormal and Social Psychology 44(1), 118-123 (1949). Accessed October 8, 2026.
  2. D. H. Dickson and I. W. Kelly, The "Barnum effect" in personality assessment: A review of the literature, Psychological Reports 57(2), 367-382 (1985). Accessed October 8, 2026.
  3. P. E. Meehl, Wanted: A good cookbook, American Psychologist 11(6), 263-272 (1956). Accessed October 8, 2026.
  4. Barnum effect, Wikipedia (for the full text of the 13 statements in Forer's sketch). Accessed October 8, 2026.
  5. A. J. Wyman and S. Vyse, Science versus the stars: A double-blind test of the validity of the NEO Five-Factor Inventory and computer-generated astrological natal charts, Journal of General Psychology 135(3), 287-300 (2008). Accessed October 8, 2026.
  6. R. Stein and A. B. Swan, Evaluating the validity of Myers-Briggs Type Indicator theory: A teaching tool and window into intuitive psychology, Social and Personality Psychology Compass 13(2), e12434 (2019). Accessed October 8, 2026.
  7. R. R. McCrae and P. T. Costa, Reinterpreting the Myers-Briggs Type Indicator from the perspective of the five-factor model of personality, Journal of Personality 57(1), 17-40 (1989). Accessed October 8, 2026.
  8. B. W. Roberts, N. R. Kuncel, R. Shiner, A. Caspi and L. R. Goldberg, The power of personality: The comparative validity of personality traits, socioeconomic status, and cognitive ability for predicting important life outcomes, Perspectives on Psychological Science 2(4), 313-345 (2007). Accessed October 8, 2026.
  9. S. Vazire, Who knows what about a person? The self-other knowledge asymmetry (SOKA) model, Journal of Personality and Social Psychology 98(2), 281-300 (2010). Accessed October 8, 2026.