← All guides

Do Personality Tests Work for Hiring? Weaker Than You'd Think.

They predict performance less than a resume screen, and nowhere near as well as a structured work sample.

The short answer

Personality tests predict job performance weakly and below other readily available methods. The best-supported trait, conscientiousness, has a validity around .20 to .22 for predicting job performance — a real correlation, but one that accounts for under 5% of the variance. The rest of the Big Five are weaker still. For context, a structured interview is typically double that, and a work sample or cognitive ability test higher again. Applicants also fake upward to a predictable degree: studies consistently find test scores shift when the same people answer in a hiring context rather than in research, particularly on conscientiousness, agreeableness and emotional stability. Legal risk is real in jurisdictions with adverse impact standards — if a personality test screens out a protected class at a higher rate and cannot demonstrate clear business necessity, it can be challenged. Yet the tests remain common, usually as one filter in a stack rather than as the primary decision point.

The validity numbers, and what they mean in selection

Meta-analyses across thousands of studies find conscientiousness has a mean corrected validity around .20 to .22 for predicting overall job performance, which translates to about 4% of variance accounted for. This is a real relationship, replicable and not zero, and it is also modest enough that many other factors matter more. Emotional stability adds a smaller contribution, and the other three Big Five traits — extraversion, agreeableness, openness — contribute variably or negligibly depending on the job.

That .20 figure is a population average, and it rises in some contexts: conscientiousness predicts better in jobs requiring sustained detail work, and extraversion matters more in roles with a high social component. The problem is that those contexts are where you could observe the behavior directly — the warehouse applicant's prior attendance record or the sales candidate's actual numbers — and both of those outpredict the trait score.

For comparison, cognitive ability tests typically yield validities of .50 or higher for complex roles, and structured behavioral interviews run .40 to .50. A work sample test in a relevant domain is often the single strongest non-GMA predictor. Personality adds something beyond those methods, but it is a small increment over a much stronger foundation, which is why it is defensible as one input among several and a poor choice as the gate.

The faking problem and why it matters differently for selection

People answering a personality test in a hiring context score higher than people answering the same instrument in research, particularly on traits employers obviously prefer. Conscientiousness, emotional stability and agreeableness all shift upward, with effect sizes in some studies exceeding half a standard deviation. This is not subtle and it is consistent enough to meta-analyse.

Two competing interpretations exist. One is that applicants are presenting their ideal selves rather than their typical ones, which means the score is aspirational more than descriptive. The other is that a high-stakes context reveals self-regulatory capacity, and the ability to present well under observation is itself predictive. The evidence leans toward the first: faked scores are generally weaker predictors of later job performance than honest scores, and forced-choice formats designed to resist faking have not eliminated the validity drop.

The practical implication is that rank-ordering candidates by personality score is less meaningful than most selection processes assume. The top scorers are often the most motivated to appear conscientious rather than the most conscientious, and distinguishing the two from a questionnaire is close to impossible.

Legal and fairness concerns

Personality tests carry real legal risk under adverse impact standards, particularly in the United States. If a test disproportionately screens out applicants from a protected class and the employer cannot demonstrate clear business necessity, it can be challenged under Title VII. Demonstrating that necessity requires showing that the test predicts job performance in this specific role, for this employer — and given the modest validities involved, that is a higher bar than many assume.

There is also a professional-ethics question separate from the law. Using a measure with a validity of .20 as a gate throws out a substantial number of people who would have succeeded, and it disproportionately affects those less familiar with the genre — first-generation applicants, older workers returning after a break, anyone who has not learned the test-taking conventions. When the same screen could be replaced with a work sample that predicts three times as strongly, choosing the personality test is a choice, and it is one worth justifying.

When and how they are defensible

As one input in a structured process, behind a resume screen and a cognitive or work-sample test, personality assessment can add value without carrying most of the weight. That is the context where the incremental validity research applies: not personality alone, but personality after the stronger predictors have already narrowed the pool.

For roles where a specific facet genuinely matters — emotional stability in high-stress safety-critical work, for example — and where that trait cannot be observed directly, a validated measure is defensible. But the specificity matters: using a general Big Five instrument when the actual requirement is narrow wastes the signal in noise.

Never use them as the primary gate, never rank-order on personality alone, and always pilot any new instrument on your own employees first to check for adverse impact before deploying it on applicants. Most selection lawsuits happen because a tool was rolled out without local validation and produced a demographic disparity the employer could not justify.

Common questions

Why do employers still use personality tests if they predict so weakly?
Partly because the tests are cheap, scalable and feel scientific, and partly because even a weak predictor is better than nothing in a context where you need to screen hundreds of applications. The problem is that they are rarely nothing — a resume screen plus a structured phone screen would outperform most personality batteries at a fraction of the cost — but the sunk cost in an existing testing contract and the inertia of the hiring process mean the tests stay in place once adopted. Some employers also report liking them for development rather than selection, which is a lower-stakes use.
Can you fail a personality test?
Not in the sense of right and wrong answers, but you can score in a pattern the employer has decided to screen out, and many applicant tracking systems filter automatically on cutoffs without a human review. The usual shape of a rejected profile is low on conscientiousness and emotional stability, though some roles filter for extraversion. If you are rejected after a personality test with no other contact, that is the likely reason.
Is it ethical to fake a personality test in a job application?
This is a common question and one you will have to resolve for yourself. The research says most applicants shift their scores upward in the desirable direction, so declining to do so disadvantages you relative to a norm that includes that behavior. It also means the scores your employer receives are partly aspirational rather than descriptive. Whether that is acceptable to you depends on your own threshold, and the fact that it is widespread does not settle the question.
What should I do if I think a personality test rejected me unfairly?
You have little recourse unless you are in a jurisdiction with strong adverse-impact protections and belong to a protected class, and even then proving the rejection was due to the test rather than some other stage is difficult. Practically, the move is to treat it as a filter you did not pass and apply elsewhere. If the rejection came after a personality test and before any human contact, the employer has told you how much individual review your application is getting, which is itself information.

Measure it on yourself

Reading about a trait and seeing your own score are different things. These assessments cover what this article describes.

Read next

Sources

  1. Barrick, M. R., & Mount, M. K. (1991). The Big Five personality dimensions and job performance: a meta-analysis. Personnel Psychology, 44(1).
  2. Schmidt, F. L., & Hunter, J. E. (1998). The validity and utility of selection methods in personnel psychology: Practical and theoretical implications. Psychological Bulletin, 124(2).
  3. Viswesvaran, C., & Ones, D. S. (1999). Meta-analyses of fakability estimates: Implications for personality measurement. Educational and Psychological Measurement, 59(2).
  4. Morgeson, F. P., Campion, M. A., Dipboye, R. L., Hollenbeck, J. R., Murphy, K., & Schmitt, N. (2007). Reconsidering the use of personality tests in selection contexts. Personnel Psychology, 60(3).

Last reviewed 2026-08-07. This article is general information about psychological measurement, not medical or psychological advice.