The Standard
Assessment · 3 min read

The personality test is not an assessment

DISC, Myers-Briggs, the Predictive Index. Widely used, defensible-sounding, and not measuring what your hiring decision needs to know.

A person at a laptop, hands spread in a shrug.

Somewhere in most hiring pipelines there's a step where the candidate answers sixty forced-choice questions and comes back as a four-letter type or a colored quadrant. It produces a clean, shareable artifact. It feels like the objective part of the process. It's the part we'd cut first.

Where these instruments came from

The type indicators trace back to mid-century frameworks built for self-understanding and counseling, not personnel selection. They were designed to help people describe themselves to themselves. Nothing about that origin story is disqualifying, and plenty of good tools got repurposed, but it matters that selection was never the design goal, because selection is a much harder problem than description.

Two things go wrong when you repurpose them anyway.

Type sorting throws away the data. Human traits are continuous. Sorting people into buckets means someone one point either side of a cutoff lands in a different category and gets a different narrative attached to them. Retest a few weeks later, in a different mood, and a meaningful share of people come back a different type. An instrument that isn't stable across a month can't be telling you something durable about a career.

Self-report is the wrong sensor in a selection context. These questionnaires ask people to describe their own tendencies. That's plausible when there's nothing at stake. In a hiring process there is something enormous at stake, and every candidate can see which answer the job wants. You are not measuring the trait. You are measuring the candidate's read on the role, which is a real skill but not the one on the scorecard.

The Barnum problem

The reports land because they're written to land. "You value both collaboration and independent focus." "You are decisive but consider input." Nearly everyone reads their profile and recognizes themselves, which is evidence of good copywriting, not good measurement. A test where every result feels accurate to its subject has no discriminating power, and discriminating power is the entire job.

The part that should worry you more

A personality profile invites a manager to build a story about a person before that person has produced anything. Once the story exists, ambiguous behavior gets read to fit it. The quiet candidate becomes "low drive" instead of "thinking." The direct one becomes "abrasive" instead of "clear." You've handed the decision a lens, and the lens was calibrated by a questionnaire.

The same tools also drift, quietly, into proxies for background, language, and culture: categories that have nothing to do with whether someone can do the work, and that a self-report instrument has no way to hold constant.

What we use instead

A work sample, scored blind against a rubric written before submissions arrive. This is the least fashionable answer in hiring and the most durable one: if you want to know whether someone can do the job, watch them do a piece of the job.

Day One™ produces an artifact you can argue with. Not "high conscientiousness," but an actual deliverable, with an actual score, against criteria you can read. When a manager disagrees with the score, they can point at the work and say why, and the rubric gets sharper. There's no version of that conversation with a four-letter type.

Where preferences do belong

Everywhere except the decision.

We ask for POP, Personal Operating Preferences, at step three of the application, before anyone has done a Day One™. How someone works best. When they do their sharpest thinking. How they want feedback delivered. It's the same category of information a personality test claims to capture. The difference is what happens to it: POP is never scored, never ranked, and never part of a pass or fail. It exists to set someone up to win and to match them with a client whose operating style won't grind against theirs.

Preferences are for onboarding. Proof is for selection. Almost all of the damage comes from swapping those two.

Common questions

As a description of someone who has nothing at stake, they're reasonable. As a selection instrument they're close to uninformative: the answers are self-reported, the candidate can see which response the job wants, and the output is a narrative rather than a measurement of the work.
It's a much better-supported model of traits than the type indicators, and it still depends on self-report in a high-stakes setting. Better psychometrics don't fix the incentive to answer strategically, and trait scores still don't tell you whether the deliverable will be good.
No. Used after the decision, as shared language for how colleagues prefer to work, it's harmless and sometimes genuinely useful. The damage comes from using it to select, screen, or explain away someone's performance before they've produced anything.
Their Day One™ work, side by side, with the rubric scores that produced it: output quality, judgment, communication, and how well they used their tools. If a manager disagrees with a score they can point at the submission and say why, which is a conversation you can't have about a four-letter type.
In POP (Personal Operating Preferences), collected at step three of the application, before Day One™. It captures how someone works best, when they think sharpest, and how they take feedback. Collecting it early is fine; scoring it never happens. It is used for matching and onboarding, never for filtering.

This is how we hire, and it's the standard we hold our own team to. The long version lives in our manifesto. If you're hiring, start here. If you're talent, start here.

One step to start

Book a call.
We build it in front of you.

Why a call instead of a form? Because you've been burned by promises before. So we don't make one — we build your Hiring Brain™ live, on the call, and you see the proof before you commit to anything.

Start hiring

Free, no commitment, and you'll see your first profiles within 72 hours.