Assessed and Misjudged: What Personality Tests in Hiring Actually Measure — and What They Miss
Somewhere between submitting your resume and receiving an offer, there is a reasonable chance you will be asked to spend thirty to sixty minutes answering questions about yourself that have nothing to do with your technical qualifications. You may be presented with a series of statements and asked how strongly you agree or disagree. You may be asked how you respond to conflict, how you prefer to receive feedback, or whether you consider yourself more analytical or more intuitive.
You are taking a personality or behavioral assessment, and depending on the organization using it, the results of that exercise may carry more weight in the hiring decision than your resume, your references, or your interview performance combined.
The use of these tools in American hiring has expanded dramatically over the past two decades. A 2023 report from the Society for Industrial and Organizational Psychology estimated that more than 75 percent of Fortune 500 companies incorporate some form of pre-employment assessment into their hiring process. The tools range from clinically validated instruments developed through decades of academic research to commercially packaged products with limited scientific grounding. And the gap between those two categories is far wider than most job seekers — or many hiring managers — realize.
What the Research Actually Supports
The scientific literature on personality and job performance is nuanced in ways that the commercial assessment industry has not always accurately represented.
The most rigorously studied framework is the Five Factor Model, commonly referred to as the Big Five, which assesses personality across five dimensions: openness to experience, conscientiousness, extraversion, agreeableness, and neuroticism. Meta-analyses spanning decades of research have found that conscientiousness — the tendency toward organization, diligence, and goal-directed behavior — demonstrates the most consistent positive correlation with job performance across a wide range of roles and industries. This finding is robust and has been replicated across diverse samples.
The predictive power of the other four dimensions is considerably more context-dependent. Extraversion, for instance, correlates positively with performance in roles that require frequent interpersonal interaction and persuasion, such as sales or client management. It correlates far less strongly — and in some studies negatively — with performance in roles that demand sustained independent concentration. An assessment that treats extraversion as a uniformly positive trait, regardless of the role being filled, is not applying the science correctly.
Beyond the Big Five, several other assessment types are commonly used in hiring contexts, with varying degrees of empirical support.
Situational Judgment Tests present candidates with workplace scenarios and ask them to select the most appropriate response from a set of options. When well-designed and validated against actual job performance data, these tools demonstrate reasonable predictive validity. Their quality is highly dependent on the rigor with which they were developed.
Cognitive ability assessments, while technically distinct from personality tests, are often bundled with them in pre-employment batteries. These are among the strongest individual predictors of job performance identified in the research literature — a finding that is both well-established and, in certain quarters, contested on equity grounds.
Typological instruments — assessments that sort candidates into discrete personality types or categories — are among the most widely used and, from a scientific standpoint, among the most problematic. The most famous of these, the Myers-Briggs Type Indicator, has been extensively criticized in the academic literature for poor test-retest reliability. A candidate taking the MBTI on two occasions separated by a few weeks will frequently receive a different type classification. Using an instrument with this characteristic to make consequential hiring decisions is difficult to defend scientifically.
How Misuse Introduces Bias and Eliminates Qualified Candidates
The consequences of deploying poorly validated assessments extend beyond inefficiency. When assessments are used as hard filters — automatically disqualifying candidates who fall below a certain score or outside a preferred profile — they can systematically exclude candidates from protected groups in ways that create legal exposure and ethical concern.
Research has documented, for example, that some assessments that penalize high scores on neuroticism may disproportionately screen out candidates with anxiety disorders who are otherwise highly qualified. Assessments that reward extraversion may disadvantage candidates from cultural backgrounds where reserved interpersonal behavior is normative rather than indicative of poor fit. And assessments with high reading demands may disadvantage candidates with learning differences or those for whom English is a second language, without any relationship to the actual requirements of the role.
The Equal Employment Opportunity Commission has issued guidance indicating that pre-employment assessments, like any other selection tool, must be demonstrably job-related and consistent with business necessity. Organizations using assessments that cannot meet this standard face potential adverse impact liability — a risk that many HR departments have not fully evaluated.
What Candidates Should Know Before They Click "Begin"
For job seekers encountering these assessments, a few strategic considerations are worth bearing in mind.
First, authenticity is generally the most effective approach, for reasons that are both ethical and practical. Many well-designed assessments include validity scales — embedded checks designed to detect response patterns that suggest the candidate is presenting a distorted or idealized self-image. Candidates who attempt to game the assessment by selecting what they believe to be the "right" answers frequently produce profiles that trigger these flags, which may result in their results being discarded or their candidacy being questioned.
Second, if you are asked to complete an assessment that feels arbitrary, poorly designed, or disconnected from the role you are applying for, that experience is itself informative. Organizations that invest in rigorous, well-validated assessment tools typically also invest in thoughtful hiring processes more broadly. The reverse correlation also tends to hold.
Third, candidates have the right to ask questions about how assessment results will be used and whether they will be shared. While employers are not legally required to share results in most contexts, a company that uses assessments transparently and is willing to discuss their role in the process is generally more trustworthy than one that treats the results as proprietary and determinative.
The Questions Employers Should Be Asking
For organizations currently using or considering personality assessments, the foundational question is not which tool is most popular or most convenient to administer. It is whether the tool has been validated against actual performance outcomes in roles similar to those being filled.
A reputable assessment vendor should be able to provide validity studies demonstrating the correlation between assessment scores and job performance metrics, along with adverse impact analyses showing how the tool performs across demographic groups. If that documentation is not readily available, the organization should treat the tool with significant skepticism.
The most effective use of personality and behavioral assessments is as one input among several — informing structured interview questions, prompting targeted conversations about work style, and surfacing dimensions of fit that a resume cannot capture. The least effective use is as an automated filter that substitutes algorithmic sorting for human judgment.
Measurement Is Not Destiny
Personality assessments, at their best, are tools for expanding the conversation between candidates and employers — a way of surfacing dimensions of fit and potential that a resume and a brief interview cannot fully reveal. At their worst, they are a mechanism for reducing complex human beings to a score, a type, or a percentile ranking that carries more authority than the evidence warrants.
The difference lies almost entirely in how they are used. For job seekers, understanding what these tools can and cannot tell an employer is the first step toward navigating them with confidence. For employers, that same understanding is the foundation of a hiring process that is both legally defensible and genuinely effective.