Open Personality Forer, 1949

We read you.

A colour, a shape, a squiggle and five taps. Then we show you a personality reading and ask how well it fits. At the end we reveal the reading’s real secret.

2 minto take part 1responses Anonymousno sign-up needed
madefor you
EXP. 059

How well can a personality reading built from a few odd questions describe you?

Loading the experiment… It needs JavaScript to run; you can still read the science box and the article below.

Take part first.

Reading the science could sway your answer. Everything unlocks once you take part, but you can read it now if you prefer.

Back to the experiment
Science box

General statements that fit almost anyone feel remarkably accurate when they are presented as being “about you”. In this experiment everyone, whatever they choose, reads exactly the same reading, word for word.

What we measure

This is an online version of Bertram Forer’s 1949 classroom demonstration. First we ask a few odd questions: a colour, a shape, the time of day you feel most productive; then we ask you to draw a squiggle on the screen and tap five times. We take a few real measurements from these (the length of the line, how often it changes direction, the rhythm of your taps) and show them on screen. Then we show you a 12-sentence personality reading. You rate it from 0 (didn’t fit at all) to 5 (describes me perfectly), then say for each sentence whether it fits you. Finally we ask how much you trust readings like this in general and whether you suspected anything while reading. The real secret is on the results screen: the reading is the same for everyone. The colour and shape you picked only changed how the screen looks; none of your answers affected the 12 sentences. We store only your rating, your sentence-by-sentence ticks, your trust and suspicion answers, and the colour, shape and time you picked; your drawing and taps never leave your device.

What the research says

Forer (1949) gave 39 students a personality test and, a week later, handed each of them a “personal” assessment. The assessments were in fact identical, and most sentences came from a newsstand astrology book. Students rated the accuracy from 0 to 5; according to Dickson and Kelly’s (1985) review the mean was 4.3, nobody rated it below 2, and only 5 students rated it below 4. Meehl (1956) named the phenomenon the “Barnum effect”, after the showman P. T. Barnum who offered “something for everyone”. Later studies showed that acceptance rises when the text is presented as made “specifically for you”, when it is favourable, and when it is said to come from a respectable-looking method (Dickson and Kelly, 1985; Snyder, Shenkel and Lowery, 1977; Furnham and Schofield, 1987).

Why it happens

Barnum statements are true of almost everyone but read as if they were personal. A handful of techniques make this work: double-headed sentences (“reserved at first, relaxed once you trust people”) cover both ends; universal experiences (“a song can take you back years”) are shared by everybody; flattering statements are accepted more readily; vague sentences (“you’re still looking for the place to show what you can do”) can’t be falsified. While reading, we match each sentence to our own memories; we quickly find a fitting example and don’t look for the ones that don’t fit. Presenting the reading as “generated from your answers” makes that matching even easier. Forer’s point was exactly this: feeling that a personality test fits you is not evidence that the test is valid.

Limitations

This is a short online version of Forer’s classroom demonstration; we use a single fixed text and there is no control group, so we don’t compare it with a genuinely individual reading. Presenting the reading as if it were computed on screen is part of the experiment; we reveal this small deception right away on the results screen. People who know about the Barnum effect, or have played before, will naturally give lower ratings, which is why we ask the suspicion question separately. The ratings are not a personality measure and shouldn’t be used to draw personal conclusions. The text exists in Turkish and English; because the wording differs slightly between the two, averages in the two languages may not be directly comparable.

4.3Mean accuracy rating of the “personal” assessment in Forer’s class (0–5)Dickson and Kelly, 1985
5 of 39Students in the same class who rated it below 4Dickson and Kelly, 1985
higher accuracy ratingsPeople told the same text was made “specifically for you”Snyder and Larson, 1972
  1. Forer, B. R. (1949). The fallacy of personal validation: A classroom demonstration of gullibility. Journal of Abnormal and Social Psychology, 44(1), 118–123. View source ↗
  2. Meehl, P. E. (1956). Wanted—a good cook-book. American Psychologist, 11(6), 263–272. View source ↗
  3. Dickson, D. H., & Kelly, I. W. (1985). The ‘Barnum effect’ in personality assessment: A review of the literature. Psychological Reports, 57(2), 367–382. View source ↗
  4. Snyder, C. R., Shenkel, R. J., & Lowery, C. R. (1977). Acceptance of personality interpretations: The “Barnum effect” and beyond. Journal of Consulting and Clinical Psychology, 45(1), 104–114. View source ↗
  5. Snyder, C. R., & Larson, G. R. (1972). A further look at student acceptance of general personality interpretations. Journal of Consulting and Clinical Psychology, 38(3), 384–388. View source ↗
  6. Furnham, A., & Schofield, S. (1987). Accepting personality test feedback: A review of the Barnum effect. Current Psychology, 6(2), 162–178. View source ↗

Can a breakfast question produce a personality reading?

Horoscopes, coffee-ground readings, “which coffee are you?” quizzes and apps promising to decode your personality all play on the same feeling: that someone or something has read you. In this experiment we want you to see that feeling with your own eyes. First we ask a few odd questions; then we show you a 12-sentence personality reading that seems to come from your answers.

You rate how accurate it is from 0 to 5, then say for each sentence whether it fits you. On the results screen we reveal the real secret: the reading is word-for-word the same for everyone. There you can see the crowd’s average rating, which sentence “fit” most often, and how you compare with Forer’s students in 1949.

Forer’s classroom

In 1949 the psychologist Bertram Forer gave 39 students in his introductory psychology class a personality test. A week later he handed each student a “personal” assessment with their name on it and asked them to rate its accuracy from 0 (poor) to 5 (perfect). The assessments were in fact identical; he had taken most of the sentences from an astrology book bought at a newsstand.

As reported by Dickson and Kelly (1985), the students’ mean rating was 4.3. Nobody rated it below 2, and only five students rated it below 4. Forer’s warning is clear: an assessment feeling accurate does not show that the method behind it can tell one person from another. A text that fits everyone equally well doesn’t describe anyone in particular.

Why “Barnum”?

Paul Meehl named the phenomenon the “Barnum effect” in 1956 while criticising personality reports in clinical psychology. The name comes from the circus owner P. T. Barnum, famous for offering “something for everyone”. Meehl’s concern was that sentences true of almost anyone were being written into reports as if they told us something about a particular patient.

Many studies in the following decades reproduced Forer’s finding. Dickson and Kelly’s (1985) review highlights two factors that raise acceptance in particular: being told the text is about you, and the text being favourable. In Snyder and Larson’s (1972) study, people told that the same text had been prepared “specifically for them” rated it as more accurate than people told it was “generally true of people”.

Anatomy of a Barnum sentence

We wrote the 12 sentences in this experiment from scratch, without using Forer’s text, following the well-known techniques. Double-headed sentences state a trait and its opposite at once: “you hang back at first, then relax once you trust people.” Almost everyone has lived both. Universal sentences describe experiences everyone shares, like a song carrying you back years. Flattering sentences (“you value deep connections”) are accepted more readily; people attribute positive traits to themselves more easily.

Vague sentences can’t be falsified: who can give a firm “no” to “you’re still looking for the place where you can show what you’re capable of”? While reading, we match each sentence with an example from our own life; we find the fitting example quickly and don’t search for the ones that don’t fit. In the solution on the results screen you can see which technique each sentence uses and what share of the crowd said it fits them.

Why does this matter?

The Barnum effect doesn’t only explain horoscopes. Some personality reports used in hiring, “scientific-looking” online quizzes and personalised advertising can all draw on the same feeling. A result fitting you well doesn’t mean it knows anything about you; the real question is how well the same result would fit everyone else.

The best defence against Barnum texts is to put your result next to other people’s. Forer’s students only realised their assessments were identical when they showed them to each other. If you share your result, remember that your friends will see the same reading; that may be the most fun part of this experiment.

FAQ

What is the Barnum effect?

It is people’s tendency to find general, vague personality descriptions that could fit anyone personal and accurate. Paul Meehl (1956) named it; the first experimental demonstration was Forer’s (1949) classroom study, which is why it is also called the “Forer effect”.

Is the reading really the same for everyone?

Yes. The colour and shape you picked only changed how the screen looks; the 12 sentences and their order are the same for everybody. Your line and taps really were measured and shown on screen, but they had no effect on the text.

I gave a high rating. Am I gullible?

No. The sentences were written so that almost everyone can find them in their own life. Forer’s psychology students gave an average of 4.3. A high rating shows you honestly compared the sentences with your life; the problem is that the text fits everyone.

Does this mean all personality tests are useless?

No. Personality scales whose validity has been studied give scores that distinguish people from one another and predict some other behaviours to a degree. The Barnum effect shows that a result feeling accurate is not evidence that a test is valid. Validity is measured by whether results show meaningful differences between people.

What data do you store?

Your accuracy rating, your “fits / doesn’t fit” mark for each sentence, how much you trust readings like this, whether you were suspicious while reading, and the colour, shape and time you picked. The line you drew and your taps never leave your device.

Discussion 0 comments

Join the discussionReading is open to everyone. Volunteer to comment and vote; it takes 20 seconds.
No comments yet.Be the first to write; a good question gets the discussion going.

Threads about this experiment

Start a new thread →
There’s no separate thread for this experiment yet. Start one to critique the method, share a paper or ask a new question.