How a Person Builds a World · Season Three: Attention, Desire, and Self-Understanding · Article 9
The arrival of a test result often brings a small sense of relief: perhaps I am not incomprehensible; I am simply a certain type. Dispersed experiences acquire a common name—why social gatherings leave me tired, why disrupted plans make me restless, why I attend to another person first in conflict. A label connects scattered parts of life into a pattern.
That clarity need not be false, but it is not automatically self-knowledge. A psychological test may be at least four things: an entertaining description, a scale investigated through psychometric research, an assessment tool used in a defined professional setting, and a language with which a person interprets a life. These uses carry very different evidential requirements and consequences.
“It sounds like me” is not sufficient evidence of validity
Bertram Forer’s 1949 experiment on the “fallacy of personal validation” became the classic source for what is now called the Barnum or Forer effect. Students completed a test and all received the same supposedly individualised description assembled from broad personality statements. They nevertheless tended to rate the description as quite accurate. A scan of the original paper is preserved by the University of Arizona.
This small classic study does not establish that everyone who recognises themselves in a test has been deceived. It demonstrates a logical limitation: the feeling of being understood does not by itself show that a description distinguishes me from other people. Broad statements that combine opposite possibilities and offer mildly favourable interpretations can readily recruit confirming episodes from memory.
“You value relationships but sometimes need solitude” or “You expect a great deal from yourself but do not always reveal your uncertainty” may fit many people. The reader supplies individual detail, making the words appear unusually specific. The person participates in completing the description while crediting the test with all of the accuracy.
Evaluating an assessment therefore requires more than asking whether its result feels flattering or familiar. What construct does it purport to measure? How were items developed? What relationship does the score have to relevant behaviour? Is it sufficiently dependable across repeated measurements? In which populations has it been examined? What degree of measurement error remains? A scale can have statistical value without supporting a strong conclusion about one individual.
Dimensions and types make different kinds of claim
Many popular assessments place people into discrete types, while a large body of personality research uses continuous dimensions. The five-factor model, for example, commonly describes broad trait dimensions such as extraversion, conscientiousness, agreeableness, neuroticism or negative emotionality, and openness. Instruments vary in quality, and applicability across languages and cultures has to be tested. The “Big Five” is not an absolute map of personality. Its dimensional language nevertheless reminds us that most people are not members of one of two natural kinds. They occupy positions along continuous distributions.
Cutting a continuous score into “introvert” and “extravert” makes two people near a boundary look categorically different. It can also make one person seem to become another type following retesting or a contextual change. Types are convenient for communication, but convenience brings a feeling of essence. “My present score is lower on this scale” becomes “this is what I am”.
Even a dependable measure must be interpreted within bounds. A personality trait is a relative tendency across time and situations, not an instruction governing every act. A generally introverted person can be an accomplished public speaker. A highly conscientious person may procrastinate on an ambiguous task. Behaviour emerges from tendency, skill, role, environment and current state. When a type substitutes for these concrete conditions, a tool of compression becomes the endpoint of explanation.
A label reorganises the evidence that follows it
After receiving a type, people begin searching life for matching episodes. Like a self-schema, the label makes certain conduct easier to notice and remember. It can provide useful language: “Long social exposure really does deplete me, so I need a boundary.” It can also become permission: “I am this type, so leadership is not for me”; “Our types are incompatible, so the relationship cannot improve.”
The second use converts a descriptive tendency into a normative limit. The assessment began by summarising responses, but its result returns to decide how the person should respond in the future. Attempts decline, counter-evidence is not produced, and the type appears increasingly essential. This is self-reinforcement, not necessarily successful prediction by the assessment.
Group identification gives a label a social life. Type-based communities offer humour, advice and belonging, which can reduce isolation. They also teach members how to translate experience into the language of the type. A difficulty involving stress, skill or relationship may be interpreted uniformly as a personality difference. A tool begins to decide which questions are available.
How test results can serve self-understanding
First, treat the result as a hypothesis rather than a verdict. Ask which concrete situations it predicts and where prediction fails. Second, distinguish score, error and category. If only four letters are provided, with no scale, norm or reliability information, there is little basis for precise authority. Third, introduce behaviour and observations by other people. A self-report scale measures how a person answers given questions under present conditions. That has value, but it is not the entirety of life.
Fourth, examine purpose. Using a test to begin discussion of work preferences or communication differences is not the same as using it to exclude a job applicant, diagnose a psychological condition or decide whether a relationship should continue. As consequences increase, so do requirements for validity, professional interpretation, challenge and alternative evidence. An entertainment test can be used lightly, but it should not quietly acquire the power of a professional judgement.
Fifth, preserve change. Personality research identifies both relative stability and change across life, roles and sustained behaviour. A difference between repeated test results may include measurement error, state fluctuation and real change. We cannot simply select whichever explanation we prefer. A test presents a dated cross-section produced under the conditions of a particular instrument.
Self-knowledge requires more than an accurate label
Even the best assessment compresses. It infers selected dimensions from selected responses and sacrifices concrete history in order to gain comparability. Self-knowledge must also understand how a tendency formed, the relationships in which it changes, the consequences it has, whether skill and environment can compensate for it, and how the person wishes to respond.
Knowing that one is prone to anxiety does not show which uncertainties produce the greatest anxiety. Knowing that one prefers solitude does not show which relationships remain worth investment. Knowing oneself as conscientious does not make every burden freely chosen. A test can bring a question to the door. It cannot complete the interpretation.
A psychological type can therefore become part of self-knowledge if it remains a tool: its measurement range is explicit; it admits error and counterexamples; it does not turn tendency into essence or description into command; and it does not monopolise a person’s authority to interpret their own life.
A label is most useful not when it finally lets us say, “This is me”, but when it supports more exact questions: “Under which conditions am I usually like this? Which different performances have I omitted? How does this tendency affect other people? Which parts should be accommodated, and which remain open to practice and change?”
Self-understanding does not end because a name has been found. A good name allows observation to begin more clearly.
Primary sources and further reading
- Bertram R. Forer, “The Fallacy of Personal Validation”, Journal of Abnormal and Social Psychology, 1949.
- Robert R. McCrae et al., research on internal consistency, retest reliability and implications for personality-scale validity, Personality and Social Psychology Review, 2011.
- Beatrice Rammstedt and Oliver P. John, “Measuring Personality in One Minute or Less”, on uses and limitations of a very brief Big Five inventory.
- Nathan W. Hudson et al., “You Have to Follow Through”, on behavioural practice and volitional personality change, 2019.
Continue reading: Explore the How a Person Builds a World series.
Discover more from Geoffrey Chen
Subscribe to get the latest posts sent to your email.