Skip to content
Seven minutes. A new way to see yourself.Get your free results →
Wellington

Guide

Personality tests for teams: what they can do and what they cannot

A personality test for teams is most useful as a shared vocabulary for how people prefer to work, not a selection or ranking tool. Conscientiousness and Extraversion relate to performance and leadership at work (Barrick & Mount, 1991; Judge et al., 2002), but those are group averages, not a way to judge one colleague. The exercise works when it stays voluntary, private and separate from any decision about anyone's role.

Last updated September 16, 2026.

What is a personality test for teams actually good for?

The honest use of a personality test in a team setting is a shared vocabulary, a set of words a group can use to talk about how its members prefer to work, communicate and handle pressure. That is a modest claim, but it is a real one. A team that can say "I do my best thinking alone before a meeting" or "I need the deadline stated plainly, not implied" is having a more useful conversation than a team guessing at the same thing from behaviour alone.

What it is not is a selection tool. Deciding who joins a team, who leads a project or who gets promoted is a different kind of decision, with different evidence requirements, and a self-report questionnaire was not built to carry that weight. Wellington is not designed for hiring, and the same caution applies to any internal decision that works like hiring: assigning roles, ranking contributors or deciding who stays on a project. Personality tests in hiring goes through why in more detail.

What do personality traits actually relate to at work?

Two findings carry most of the weight when a trait shows up in a work context. Conscientiousness, the trait behind planning, reliability and follow-through, relates to job performance across every occupational group studied, from professionals and police officers to managers, sales staff and skilled workers (Barrick & Mount, 1991). Extraversion is the most consistent trait correlate of leadership across settings, at a corrected correlation of .31 (Judge et al., 2002), though the criteria in that review mix leader emergence with rated effectiveness, so it is not a finding that extraverts do the job better.

Both findings deserve the same caution in a team setting that they deserve in hiring. They describe how a large group tends to behave, on average, across many teams and many jobs. They do not describe your team, and they do not tell you which named person on it will perform best or lead best. A trait exercise that treats a group-level correlation as a fact about one colleague has already gone further than the evidence supports. The Big Five and your career covers what these findings do and do not support for an individual.

Why do colleagues often see us differently than we see ourselves?

People are not equally good judges of every part of their own personality, and that gap is exactly what makes a team exercise worth doing. Self-report is more accurate for low-visibility, internal states, such as how anxious or emotionally steady someone tends to feel, while people who know someone well judge more evaluative, outward-facing traits, such as how capable or organised they seem, more accurately than the person judges themselves (Vazire, 2010). Extraversion is the trait where self and others judged about equally accurately in that study.

That asymmetry runs through teams constantly and mostly goes unspoken. A colleague's self-description of their own working style and a teammate's read of the same person can both be right, about different things. Formally, observer ratings from people who know someone predict outcomes such as job performance about as well as self-ratings, and better for some outcomes (Connelly & Ones, 2010), which is the measured version of a pattern any long-running team has already noticed informally. A structured exercise that puts a person's own answers next to how teammates would describe them, done with care and without ranking anyone, gives a team language for a gap that is usually there whether or not anyone names it.

Why do type-based team tools over-read their own results?

Many team-building tools sort people into a small number of named types rather than reporting a trait as a score. That format has been examined directly for the best-known instrument of this kind, and the conclusion is a caution rather than an endorsement: its four-letter formula does not support the inferences commonly drawn from it, and considerable care is warranted before leaning on it for decisions in a workplace setting (Pittenger, 2005).

The underlying problem is structural. Each trait a type label draws on is a continuous dimension rather than a set of distinct types, and across a large sample scores are roughly bell-shaped, with most people near the middle and few at either end (McCrae & Costa, 1989; Goldberg, 1993). There is no natural line on that curve, which is why sorting people into a handful of boxes throws away real information, and why a colleague sitting close to a boundary can land in a different box on a retake without having changed at all. A team that reorganises itself around four or sixteen fixed types is reorganising around a boundary that was never there in the data.

What makes a good team personality exercise, and what does not?

The difference between a useful session and a wasted one usually comes down to what the group does with the result afterwards.

Good and poor uses of a team personality exercise
Good use of a team exercisePoor use of the same result
Giving the team words for different working styles, so preferences can be named instead of guessed atDeciding who leads the project or owns the next task based on a score
Prompting a conversation about how the team prefers to receive feedback or handle conflictRanking teammates against each other on any trait
Helping a manager ask better questions in a one-to-oneTreating a score as an explanation for a specific past disagreement
Normalising that colleagues can read the same person differently (Vazire, 2010; Connelly & Ones, 2010)Announcing someone's scores to the group without asking them first
Revisiting the exercise occasionally as a light check-inFiling the result anywhere near a performance review
  • Treat every result as a starting point for a conversation, not a verdict on a colleague.
  • Keep the exercise separate, in writing and in practice, from any decision about roles, pay or standing on the team.

What ground rules keep a team exercise honest?

A few simple rules protect a team personality exercise from becoming the thing it should never be.

  • Voluntary. Nobody should feel their standing on the team depends on taking part.
  • Private scores. Each person decides what, if anything, they share with the group, and a manager should not collect results directly.
  • No ranking. There is no better or worse trait profile for being a good colleague, only different working styles.
  • No decisions from it. Roles, assignments, pay and promotion are decided on evidence suited to that purpose, never on a self-report exercise built for reflection.

Held to those rules, a team exercise stays what it can honestly be: a way to talk about working styles. Dropped, and it starts to function like an unofficial, unvalidated hiring test aimed at people who already have the job. Self-report vs observer report goes further into why self and other perspectives on the same person can both be trustworthy and still disagree.

How does Wellington measure this?

Wellington is a science-based personality assessment from Therabot Labs LLC. It measures personality as continuous traits, built on the Big Five and HEXACO models, and writes the results back as a personal report rather than a type.

Wellington is built for the person taking it, one Snapshot at a time, not for a manager or a team to administer. Each person takes their own free Snapshot, 60 questions, about 7 minutes, scoring 6 dimensions: Extraversion, Agreeableness, Conscientiousness, Emotional Stability, Openness to Experience, Honesty-Humility. The first ten questions are answered before anything is asked of you; after question ten you sign in with an email so your answers are saved and your results can be shown. There is no card.

Every trait is reported as a percentile against the reference sample. The three bands (Quiet below the 30th percentile, Balanced to the 70th, High above) and the Portrait names exist to make the result easier to talk about; the percentile is the measurement, and it is always shown alongside. Norms are provisional until the norming sample is complete.

Sharing is a choice each person makes for themselves, not a team-wide export. A Portrait card can be shared if a colleague wants to bring their result into a team conversation, but nothing leaves an account unless the person chooses to share it. The Wellington Membership, $4.99 a month, adds the Portrait, 336 further questions across 4 chapters, to reach 90 traits, then continues toward all 255 traits. Wellington is a wellness tool for reflection, not a medical or psychological diagnosis, and it does not replace care from a qualified professional. It is not designed for hiring, selection or any decision about someone's standing on a team, and it should not be used to assess anyone other than yourself. You can export or delete your data at any time.

Questions people ask

Should a manager use a personality test to decide who leads a project?
No. Extraversion is the trait most consistently linked to leadership (Judge et al., 2002), but that is a group average across many samples, not a fact about one named person. Deciding who leads is a decision with fairness and performance stakes that a self-report questionnaire was never built to carry, and Wellington is not designed for that use.
Is it normal for teammates to describe someone differently than they describe themselves?
Yes, and it is often informative rather than a disagreement to resolve. Self-report is stronger for low-visibility traits such as how anxious someone feels, while people who know someone well tend to judge more evaluative traits more accurately (Vazire, 2010), and observer ratings predict outcomes like job performance about as well as self-ratings, and better for some (Connelly & Ones, 2010).
Are type-based team tools, like a four-letter type, reliable for building a team?
Not on the evidence. The best-studied type instrument's four-letter formula does not support the inferences commonly drawn from it (Pittenger, 2005), and its scales measure continuous dimensions rather than distinct types (McCrae & Costa, 1989). A colleague near a boundary can land in a different type on a retake without changing at all.
Should results from a team personality exercise be shared with the whole group?
Only if each person chooses to share their own. A useful exercise keeps scores private by default and lets each colleague decide what to bring into the room. Nothing should be shared, filed or discussed without that person's choice, and results should never be collected by a manager or used near a review.
Can a personality test for teams improve how a team works together?
It can, as a shared vocabulary for working styles rather than a verdict on anyone. Use it to prompt conversation about feedback, pace and communication, and keep it away from ranking or assigning people, which the trait evidence was never built to support. There is no study showing that a team exercise by itself improves team performance, so treat it as a conversation starter.

Sources

Peer-reviewed sources for the claims above. Wellington's own reliability figures will be published once the norming sample is complete.

  1. Barrick, M. R., & Mount, M. K. (1991). The Big Five personality dimensions and job performance: A meta-analysis. Personnel Psychology, 44(1), 1–26.
  2. Judge, T. A., Bono, J. E., Ilies, R., & Gerhardt, M. W. (2002). Personality and leadership: A qualitative and quantitative review. Journal of Applied Psychology, 87(4), 765–780.
  3. Vazire, S. (2010). Who knows what about a person? The self-other knowledge asymmetry (SOKA) model. Journal of Personality and Social Psychology, 98(2), 281–300.
  4. Connelly, B. S., & Ones, D. S. (2010). An other perspective on personality: Meta-analytic integration of observers' accuracy and predictive validity. Psychological Bulletin, 136(6), 1092–1122.
  5. Pittenger, D. J. (2005). Cautionary comments regarding the Myers-Briggs Type Indicator. Consulting Psychology Journal: Practice and Research, 57(3), 210–221.
  6. McCrae, R. R., & Costa, P. T. (1989). Reinterpreting the Myers-Briggs Type Indicator from the perspective of the five-factor model of personality. Journal of Personality, 57(1), 17–40.
  7. Goldberg, L. R. (1993). The structure of phenotypic personality traits. American Psychologist, 48(1), 26–34.

Read next

See your own traits, not a type.

The free Snapshot takes about seven minutes and gives you your personality card and five dimensions. No credit card, no type.

One trait a week

Not ready to take the test? Read one trait a week.

A real page from the report, one practice attached, every week. It is the easiest way to see whether Wellington reads people the way you think it should.