ReachLink is now hiring licensed therapists. Apply to join the current cohort before September 30. Apply now →

What the Myers Briggs Actually Measures About You

TestsSeptember 24, 202616 min read
What the Myers Briggs Actually Measures About You

The Myers-Briggs Type Indicator measures self-reported preferences across four dichotomies, sorting people into one of 16 personality types based on Jungian theory, yet research shows it lacks strong test-retest reliability and predictive validity, making licensed therapy a more reliable path to genuine self-understanding.

Ever taken the Myers Briggs twice and landed on a different type each time? That's not a fluke, it's built into how the test works. Here's what those four letters really measure, what they don't, and why your need for self-understanding deserves more than a label.

What the Myers-Briggs Type Indicator actually measures

The Myers-Briggs Type Indicator (MBTI) is a self-report questionnaire. You answer a series of forced-choice items, meaning you pick between two statements that each describe a way of thinking or behaving, and there is no option to say both apply or neither does. Nothing about the process involves an outside observer rating you or a clinician making a judgment. It is closer to a structured survey of your own preferences than to any kind of assessment a mental health professional would use to evaluate a condition.

What does the Myers-Briggs Type Indicator measure?

The MBTI sorts your answers into four separate dichotomies, each framed as a pair of opposite preferences. These are extraversion versus introversion, sensing versus intuition, thinking versus feeling, and judging versus perceiving. Each dichotomy is meant to capture a general tendency, not a fixed trait you either have or lack entirely. Put together, the four letters you land on make up your reported type.

The four dichotomies and how they are scored

Your raw score on each dichotomy actually falls somewhere on a continuous scale, meaning most people land somewhere in the middle rather than at either extreme. The instrument then converts that continuous score into a single letter, so someone who scores nearly 50-50 between thinking and feeling gets sorted into the same category as someone who scores heavily toward one side. That conversion is where a lot of the disagreement about the tool starts, though the mechanics of why retesting produces different results are a separate question. The letter you receive tells you which side of the midpoint you fell on, not how strongly.

What a four-letter type is meant to describe

Combining the four letters produces one of 16 personality types, each given a label like ENFP or ISTJ. These types are described as preferences for how you direct energy, take in information, make decisions, and organize your life, not as abilities, skills, or guarantees of how well you will perform at anything. The MBTI does not claim to measure mental health, pathology, competence, or job success. That distinction matters: a type describes a preference, while something like the diagnostic criteria behind personality disorders describes a pattern of functioning that causes real distress or impairment, and the two are not measuring the same kind of thing at all.

Type dynamics and the dominant function claim

Beyond the four letters, MBTI theory includes the idea of type dynamics, which proposes that each type relies on an internal hierarchy of four mental functions, with one function operating as dominant and the others playing supporting roles. This dominant function claim suggests that your type is not just four independent preferences but a structured system with a lead process guiding the rest. It is a more elaborate layer than most people encounter when they first take the test, and it moves well past simple description into a theory about how the mind is organized internally.

Where the framework came from, and why that matters

The MBTI history begins not with a questionnaire but with a book. Carl Jung published his theory of psychological types in 1921, describing patterns he had observed in patients and in himself: whether a person’s energy orients outward or inward, and how they take in and judge information. Carl Jung psychological types was a work of philosophy and clinical interpretation, built from case observation rather than measurement. Jung was not trying to build a test. He was trying to describe a structure he believed he saw in the mind.

Katharine Cook Briggs read Jung’s work and became fascinated by it. Her daughter, Isabel Briggs Myers, later turned that fascination into a questionnaire meant to sort people into Jung’s categories. Neither woman had formal training in psychometrics, the discipline concerned with how psychological instruments are constructed and validated. They worked largely outside academic psychology, refining the tool through decades of independent effort rather than university research.

The instrument found its footing in offices rather than clinics. Businesses adopted it for hiring, team building, and career guidance long before psychology departments took much interest in it. That path mattered, because it meant the tool spread and calcified as a workplace habit while academic personality science moved somewhere else. Researchers in that period increasingly favored factor-analytic models, built by statistically grouping traits that actually correlate across large samples, a different starting point than Jung’s typology.

None of this makes the instrument automatically wrong. Origin is not the same as evidence. It does explain a real gap: the theory existed first, built by observation, and the questions about whether it holds up came only after it had already spread far beyond where it started.

Why people get a different type on retest

A lot of people take the Myers-Briggs more than once and land on a different set of letters the second time. That is not random bad luck. It comes from how the test is built and how it turns a number into a label.

What reliability means for a personality test

Reliability means an instrument gives a stable answer when nothing about the person has changed. If you take a test on Monday and again on Friday, and you have not changed in any meaningful way, a reliable test should hand back roughly the same result. Mbti test-retest reliability is the specific question of whether the same person gets the same type letters back on a second try. When the answer is often no, the problem lies in the test itself, not in the person’s stability.

The midpoint problem, explained

Each dichotomy on the Myers-Briggs, thinking versus feeling for example, is scored on a continuous scale before it gets chopped into a single letter. A dimensional score would report something like slightly toward thinking. A categorical label reports T with the same confidence it reports a strong, lopsided preference. People who land near the middle of that scale are the ones most likely to flip letters, and midpoint scores are common rather than rare. A flip changes the label, the type description, and any advice attached to it, even though the underlying score barely moved.

What the retest studies looked at

Reliability research on the instrument has looked directly at how often people change type across repeated testing. Work described in cautionary comments regarding the Myers-Briggs Type Indicator by Pittenger examined this pattern, as did research by McCarley and Carskadon on how consistently respondents landed in the same category across retest intervals. A related psychometric analysis of the MBTI’s scoring limitations ties this instability back to the same source: continuous scores forced into categories at a fixed cut point. Mood, recent events, and the setting in which someone takes the test can all nudge a borderline answer just enough to cross that line, which is often the honest answer to why does my mbti type change between sittings.

Validity, and what the test can and cannot predict

Reliability asks whether you get the same answer twice. Validity asks something different: does the instrument measure what it claims to measure, and does that measurement tell you anything useful. This is the more serious question for the Myers-Briggs, and it is the one where mbti validity runs into trouble. A description can feel exactly right and still fail this test, because accuracy of feeling and accuracy of measurement are not the same thing.

Do the categories describe real divisions between people?

The theory claims that people fall into one of two camps on each dichotomy: extraversion or introversion, thinking or feeling, and so on. If that were true, you would expect trait scores to cluster into two separate humps, one for each side. Instead, research on personality trait structure finds that traits distribute the way height or blood pressure does, bunched around the middle with fewer people at the extremes. That pattern undercuts the idea of a genuine type category and points instead to a continuum, which is a different claim than the one the instrument is built on.

The four dichotomies are also supposed to be independent, so that any of the 16 combinations is equally plausible. In practice they are not fully independent of each other, which weakens the claim that each of the 16 types is a distinct, freestanding combination rather than a handful of overlapping tendencies.

Why do psychologists remain skeptical of the Myers-Briggs? What are the criticisms of the Myers-Briggs Type Indicator?

A review that applied a unified framework for test validity, requiring multiple independent sources of evidence rather than a single validation study, concluded that there is insufficient evidence to support the tenets and practical claims made about the instrument. That is the core of the skepticism: not that the instrument feels wrong to take, but that the case for what it measures and what it predicts has not been built to the standard modern validity research expects.

What type does not predict at work

Is Myers Briggs scientifically valid as a predictor of job performance, team fit, or career success? The evidence has generally not supported using type for hiring, placement, or team assignment decisions. Notably, the instrument’s own publisher discourages its use in hiring and screening, which is itself a statement about the limits of what type can predict.

Why an accurate-feeling description is not evidence

Myers Briggs accuracy, in the sense of feeling personally true, and predictive validity are two separate questions, and one does not answer the other.

Does the theory behind the types hold up to scientific testing?

A scientific claim earns that label by specifying, in advance, what result would prove it wrong. Judged against that standard, the theory behind the Myers-Briggs types gets a mixed verdict rather than a single grade. Some of its claims can be tested and have not held up. Others are built in a way that makes testing impossible from the start.

Is MBTI pseudoscience or just an untestable theory

The dichotomy claim is the clearest example of something testable. It predicts that people cluster into two distinct groups on each dimension, introvert or extravert, thinker or feeler, with few people sitting in the middle. That is a real prediction, and it can be checked against how traits actually distribute in the population. An analysis of the theoretical validity behind the Myers-Briggs Type Indicator found that this is exactly where the theory fails: traits show up as smooth, continuous curves rather than two separate peaks, which is the opposite of what a clean dichotomy would predict.

Type dynamics and the dominant function hierarchy sit in different territory. There is no independent way to measure which function sits highest in someone’s hierarchy apart from the four letters that already assume it. Without a separate measurement, there is no result that could disconfirm the hierarchy. That is the heart of mbti falsifiability as a problem: a claim with no possible failure condition has not been tested, it has been assumed.

Curious about something here?

Ask your favorite AI about this article

Type descriptions compound this. They are written broadly enough that almost any behavior can be read as fitting, which means the descriptions resist being wrong. The same review points to this as one of the theory’s structural flaws: its practical popularity and its resistance to disproof exist side by side.

Myers-Briggs, Big Five, and HEXACO side by side

Psychologists who study personality tend to reach for different tools than the MBTI. The two most commonly cited alternatives, the Big Five and HEXACO, were built from different starting points and report their results in a different shape. Comparing them side by side shows what a difference between Myers Briggs and Big Five actually looks like in practice, not just in theory.

How the Big Five was built differently

The Big Five model did not start with a theory about how personality works. Researchers analyzed the words people use to describe each other, looked at which words clustered together across many samples, and let five factors emerge from that pattern: openness, conscientiousness, extraversion, agreeableness, and neuroticism. Extraversion shows up in both frameworks, but the Big Five reports it as a position on a continuous scale rather than a single letter. A person who scores in the middle is reported as exactly that, in the middle, rather than sorted into one side or the other. Conscientiousness and emotional stability in particular have been studied for decades as predictors of outcomes like job performance, and a comparative review of personality and other predictors of life outcomes found these continuous trait measures carry real predictive value.

What HEXACO adds

HEXACO keeps much of the Big Five’s structure but adds a sixth dimension: honesty-humility. This factor covers tendencies like sincerity, fairness, and avoiding manipulation of others. Adding that sixth dimension also reorganizes how some of the other traits, particularly agreeableness and emotionality, get separated out. The result is still dimensional and still self-report, but it captures a slice of behavior the original five factors were not designed to isolate.

Comparison table: model, output, evidence base, and limitations

  • MBTI: reports four-letter types built from dichotomies; derived from Jungian theory; treats preferences as stable categories; commonly used for self-reflection and team exercises; limited predictive evidence.
  • Big Five: reports scores on five continuous dimensions; derived from statistical clustering of trait language; scores can shift somewhat over time; used in research and organizational settings; still self-report and describes tendencies, not fixed outcomes.
  • HEXACO: reports scores on six continuous dimensions; derived by extending Big Five methods; same self-report limitations as the Big Five; used in research on ethical and prosocial behavior.

None of these models predicts a fixed future. A high or low score describes a tendency, not a destiny, and every one of them relies on the person answering honestly about themselves.

Why the four letters feel so accurate

Type descriptions are written to fit. Most of the language is positive, broad, and phrased so that almost anyone reading it can find themselves somewhere in the paragraph. Psychologists call this the Barnum or Forer effect, named for the observation that vague, flattering statements feel personally accurate even when they would apply to nearly anyone. This is a large part of why people love MBTI results: the letters arrive attached to a description that reads like it was written for you specifically, when it was written to work for almost everyone.

Being handed a coherent, flattering account of yourself feels good, especially if self-understanding has felt hard to reach. A four-letter code also gives people something concrete to hand each other. Saying you are an INFJ or an ESTP creates instant shorthand, and shorthand builds a sense of belonging that has nothing to do with whether the test meets scientific standards. That social usefulness is real, even separate from the ongoing debate over whether MBTI scores measure anything meaningful.

The trouble starts when the label stops describing a tendency and starts acting like a ceiling. Over-identification can look like treating a letter as a fixed trait you cannot work around, using it to explain away a conflict instead of looking at what actually happened, or turning down an opportunity because it feels off-type. There is a real gap between “this is how I tend to operate” and “this is what I am and cannot change,” and the four letters make it easy to slide from one into the other without noticing. Once that happens, the label can quietly take the place of asking what is actually going on in a specific situation, which is usually a more useful question than which letters you were assigned.

Using type language without letting it box you in

A letter can still be useful if you treat it as a hypothesis rather than a verdict. The test is simple: does it match what you actually did last week, not what you feel like in the abstract.

Turn the letter back into a behavior

Instead of saying “I’m an introvert,” name the specific settings that drain you and the ones that recharge you. Maybe large meetings wear you out but a two-person work session doesn’t. Maybe a full day of client calls leaves you flat, but an hour alone at the end of it brings you back. The behavior is checkable against your week. The label isn’t.

When a personality label is the wrong tool

Watch for the moment a type gets used to close a conversation instead of open one, whether you’re doing it to yourself or a team is doing it in a meeting. “That’s just how ENTJs are” ends the discussion instead of starting one. Type also has no way to assess distress. If what’s underneath the question is anxiety, a low mood that won’t lift, burnout, or a relationship stuck on the same fight, no personality framework substitutes for talking it through with someone. That kind of work, and relationship patterns that keep repeating, are also part of what couples therapy is built to address when both people are involved.

Your own data beats a category

Among personality test alternatives, the simplest is also the most specific: a mood log or plain journal kept over several weeks. It shows you your actual patterns, not four letters standing in for them. If self-understanding without labels is the actual goal, a working alliance with a licensed therapist looks at what you did, felt, and said, not which category you fit. You can create a free ReachLink account and browse licensed therapists at your own pace, with no commitment, if that’s the kind of clarity you’re after. That kind of looking, through psychotherapy, starts from your own record, which a category can’t give you.

Wanting a clear answer for who you are makes complete sense

The pull toward a four-letter label is not shallow curiosity. It comes from a real wish to feel understood, to have language for the parts of you that feel hard to explain to other people, maybe even to yourself. That wish deserves more than a quiz can give it, and the skepticism from psychologists does not erase the real relief people find in feeling named. Both things can be true at once: the tool may be imperfect, and your need to understand yourself is not.

If you have been using personality frameworks to make sense of patterns that feel bigger than a type, a licensed therapist can help you look at them with more depth and care than any test allows. You can begin with a free assessment at ReachLink, at your own pace and with no commitment, and let a care coordinator match you with someone suited to what you are working through: begin with a free assessment at ReachLink.


FAQ

  • Why do I get a different Myers-Briggs type every time I take it?

    The Myers-Briggs converts continuous scores into one of two letter categories for each of its four dimensions, so someone who scores near the middle of the thinking-versus-feeling scale gets sorted the same way as someone with a strong preference in either direction. Because many people naturally score close to the midpoint on at least one dimension, small shifts in mood, recent events, or even the setting where they take the test can push a borderline score across the cutoff and produce a different letter. Research on the instrument's reliability has found that a notable portion of people receive a different type on retesting, even within a short period of time. If your results keep changing, it most likely reflects where your score sits on the scale, not instability in who you are as a person.

  • Can therapy help if I feel like I'm using personality labels to avoid dealing with real problems?

    Yes, and this is one of the more common patterns a therapist can help you examine. Relying on a personality type to explain away conflict, deflect feedback, or avoid change can become a way of closing conversations that would benefit from staying open. In therapy, a licensed therapist can help you look at the specific behaviors, patterns, and situations that underlie the labels you have been reaching for, rather than working with the label itself. Approaches like cognitive behavioral therapy (CBT) are well-suited to identifying the thoughts and habits that keep you stuck, regardless of what letters a personality test assigned you.

  • Why does my Myers-Briggs result feel so accurate if psychologists say the test isn't that reliable?

    The feeling of accuracy and the scientific validity of a test are two separate things, and the Myers-Briggs is a clear example of why they do not always match. Type descriptions are written using broad, positive language that is easy to see yourself in, a well-documented phenomenon psychologists call the Barnum or Forer effect, where vague and flattering statements feel personally true even when they would apply to almost anyone who reads them. That sense of recognition is real and worth taking seriously, but it tells you more about how description works than it does about whether the test is measuring something precise about you.

  • I've been thinking about starting therapy to understand myself better - how does ReachLink work and where do I begin?

    ReachLink connects you with licensed therapists through a free assessment and a human care coordinator who reviews your situation and makes a thoughtful match, rather than using an algorithm to pair you with someone automatically. After you complete your free assessment, your care coordinator considers what you are working through, your preferences, and your availability to find a therapist who is a good fit. From there, you meet with your therapist through ReachLink's telehealth platform, so sessions happen wherever you are most comfortable. You can start with no commitment by creating a free account at ReachLink and completing the initial assessment at your own pace.

  • Are personality tests like the Myers-Briggs ever actually useful, or should I just skip them?

    Personality frameworks can be a useful starting point for self-reflection, as long as you treat the result as a loose hypothesis rather than a fixed description of who you are. The risk comes when a label starts doing the thinking for you, steering decisions, closing off possibilities, or standing in for a more honest look at what is actually happening in a specific situation. If you want a more grounded picture of your own patterns, keeping a simple journal or mood log for a few weeks gives you concrete, checkable data rather than a category. And if the patterns you are noticing feel bigger than a type can explain, talking with a licensed therapist is a much more specific and effective tool for that kind of self-understanding.

Have a question about this topic?

Type your question and we'll send it to the AI assistant of your choice.

Your question will be sent to an external AI assistant. If you're going through a crisis, please reach out to the 988 Suicide and Crisis Lifeline (call or text 988).

Share this article
Take the First Step

Get Real Support.
See Real Results.

Join thousands who have found specialized therapy that truly understands their health journey. Start today — it takes less than 5 minutes.

No referral needed · Most insurance accepted · Start within 48 hours