In brief

For a specific job competency, a job-analysis-based structured interview is usually the more direct of these two methods because it asks about relevant behavior and scores answers against shared criteria. A personality report describes measured tendencies; it can guide reflection or follow-up questions, but it does not prove that an individual can perform a particular task. Use a report for selection only when evidence supports that exact instrument, interpretation, population, and use.

What does each method actually measure?

If the decision is whether a person can demonstrate a specific job competency, a structured interview is usually the more direct starting point. It asks about behavior in relevant past situations or a defined hypothetical, then applies common questions and scoring standards. A personality report summarizes answers to items as tendencies on the instrument’s scales. It may help someone reflect on how they usually approach work, but a tendency score does not itself show that the person can complete a particular task. The right choice therefore depends on the claim you need to support: demonstrated behavior, typical self-described pattern, or some other outcome.

The U.S. Office of Personnel Management describes structured interviews as methods designed to measure job-related competencies through systematic questions about past behavior or proposed behavior in hypothetical situations. Its guidance says candidates receive predetermined questions and responses are assessed using the same scale and standards. That consistency makes answers easier to compare, but does not guarantee accuracy. A badly chosen competency or vague rating guide can produce orderly scores about the wrong thing. A personality report has a different measurement chain: item responses are scored into a scale, then interpreted using the instrument’s construct definition, evidence, and any stated comparison group. Not every report uses norms, and a polished narrative does not prove that an interpretation is valid for employment.

A competency is a defined capability relevant to work, expressed in actions or results. A personality tendency is a recurring way of thinking, feeling, or behaving that a measure attempts to describe. Validity concerns whether evidence supports a particular interpretation and use of scores. It is not a general badge attached to a test. Thus, “this person says they like planning” and “this person can coordinate a project to deadline” are separate claims. One may suggest a question about the other; it cannot stand in for it without evidence connecting the scale to that inference and purpose.

Sources: Structured Interviews; Job Analysis

Start with the competency, not the report label

Before comparing methods, write down the work behavior the decision requires. “Organized,” “leadership,” and “good under pressure” are labels, not yet observable standards. For planning, a specific competency might be maintaining a workable schedule when dependencies change: identify milestones, surface a blocked task, communicate a revised sequence, and follow through. These behaviors can be discussed in an interview and, where suitable, observed in a job-like exercise. They differ from preferring tidy lists or feeling comfortable with advance plans.

OPM calls job analysis the foundation for assessment and selection. It examines tasks, the competencies needed for those tasks, and the links between them. In practical terms, start with critical work episodes: what happens, what choices matter, and what effective performance looks like. Then decide whether the target is knowledge, a practiced skill, a behavior under specified conditions, or a general tendency. A report’s broad scale may be relevant to the last category. An interview question about a real dependency failure may elicit evidence about the first three. If a report says little about the target behavior, the report label cannot bridge the gap.

Professional selection principles similarly emphasize a reasoned connection between work requirements and assessment content. The 2018 SIOP Principles, adopted as APA policy, discuss work analysis and matching assessment procedures to specified requirements. This is guidance for building and evaluating selection procedures, not a finding that one format always wins. The useful reader question is: what would count as evidence, and does this method actually produce it? If the role requires responding to a late supplier change, an account of how a person handled a comparable change is closer to the target than an abstract preference for structure. The former is still a report of behavior and needs careful scoring; it is simply more directly tied to the decision.

Sources: Job Analysis; Principles for the Validation and Use of Personnel Selection Procedures, Fifth Edition

Why structure changes what an interview can support

Structure makes an interview more comparable when it fixes the core questions, gives candidates similar opportunity to answer, and uses explicit criteria. OPM’s guidance recommends questions about relevant past behavior or hypothetical situations and common rating standards. A behavioral prompt might ask how the candidate handled a project plan after a key input arrived late. A situational prompt could ask what steps they would take if a deadline became threatened by a new dependency. The interviewer can use planned probes to clarify the person’s actions, reasoning, and outcome rather than improvising different standards for each candidate.

A meta-analysis by McDaniel and colleagues reviewed 245 validity coefficients from 86,311 people and found interview validity varied with interview content, degree of structure, format, and the criterion being predicted. The analysis is substantial but dates from 1994 and draws on the studies then available; it is not a current head-to-head test of every personality instrument against every interview. Its important implication is conditional: an interview’s usefulness changes with what is asked, how the procedure is run, and what outcome counts as success. Structure is a design feature, not a magic word.

A separate review of 338 ratings from 47 interview studies examined the constructs interviews measured. Huffcutt and colleagues reported that personality and applied social skills were frequently rated and suggested that structured interviews’ higher validity may partly reflect attention to constructs more strongly related to job performance. This complicates the claim that standardization alone explains better evidence. The content matters too. A fixed set of irrelevant questions, or a rubric that rewards confidence instead of planning behavior, remains weak evidence. Conversely, an interview may be structured to assess interpersonal judgment rather than task knowledge, if that is the actual competency and the method has a defensible design.

When reviewing an interview, ask whether the questions came from work requirements, whether all candidates are assessed on comparable prompts, and whether the rating anchors distinguish levels of relevant evidence. A score should reflect the answer against those criteria, not general likability or resemblance to the interviewer. Multiple trained raters can reduce dependence on one person’s impression, though they do not remove judgment or error. For a precise capability, direct follow-up on actions, constraints, tradeoffs, and results is more informative than an unstructured conversation that drifts wherever the interviewer’s curiosity leads.

Sources: Structured Interviews; Comprehensive Review and Meta-Analysis of the Validity of Interviews; Identification and Meta-Analytic Assessment of Psychological Constructs Measured in Employment Interviews

What personality evidence can—and cannot—add

Personality evidence should not be dismissed wholesale. Tett, Jackson, and Rothstein’s 1991 meta-analysis reviewed 494 studies, identifying usable results for 97 independent samples totaling 13,521 people. It reported that associations between personality measures and job performance differed by research strategy and were higher in studies that explicitly used job analysis to select measures. The study therefore supports a narrower point: personality measures can be relevant to job outcomes, and deliberate matching to work matters. It does not establish that every current report predicts every competency, or that an individual with a particular score can or cannot perform a task.

The distinction is between group-level evidence and an individual decision. A correlation across a sample describes how two measured variables tend to vary in that studied context. It does not identify the cause of a particular person’s behavior, establish their skill, or tell you exactly how they will act in one future situation. The relation may depend on the scale, job family, outcome definition, sample, and method. A meta-analysis can summarize patterns across studies, but variation and measurement choices still affect the conclusion. That is why a broad statement such as “conscientious people perform well” cannot substitute for evidence that a given measure supports an inference about coordinating a particular project.

Barrick and Mount’s 2001 PubMed-indexed meta-analysis explicitly revisited measurement choices in the Big Five and job-performance literature. Its accessible abstract argues that earlier syntheses included data not based on actual Big Five measures and presents a more critical interpretation of the relationship. That is useful counterevidence to simplistic claims of universal personality prediction. It does not mean personality is irrelevant; it shows that construct labels and score sources must be examined rather than treated as interchangeable. Older syntheses also predate current instruments and work contexts, so they should frame questions, not validate a contemporary product by association.

A report can be useful in development when its language prompts a testable question: Does the person tend to start planning early, or do they plan once priorities are clear? What happens when a routine changes? A reader can compare that hypothesis with concrete episodes, feedback, and constraints. For employment selection, the same report needs evidence for the exact interpretation and intended use. Reliability, meaning score consistency, is not validity: consistent measurement of a broad tendency does not establish that the score measures project scheduling skill. Nor does narrative specificity create accuracy. Use the report to guide inquiry unless the instrument-specific evidence supports a stronger use.

Sources: Personality Measures as Predictors of Job Performance: A Meta-Analytic Review; The Big Five Personality Dimensions and Job Performance: A Meta-Analysis; Principles for the Validation and Use of Personnel Selection Procedures, Fifth Edition

When could a personality report be relevant to selection?

A personality measure could contribute to selection when the intended construct is relevant to the work and evidence supports interpreting the specific instrument’s scores for that use and population. The fact that personality dimensions have sometimes related to performance is not enough. A selection decision requires a clear account of the job requirement, the score’s meaning, the criterion it is expected to relate to, and the evidence linking them. The appropriate evidence may differ across jobs and instruments. A report designed for personal reflection should not be silently repurposed as a hiring screen because its descriptions sound plausible.

The SIOP Principles caution against deciding validity questions from a construct label alone; work analysis and the relation between procedure content and job requirements matter. OPM likewise links selection tools to job analysis and recommends documenting tasks, competencies, and tool content. Together these sources support a method-matching rule, not an automatic ban on personality measures or a blanket endorsement. If the job genuinely requires a stable tendency and an instrument has appropriate evidence for the proposed score use, a validated personality measure may add evidence alongside other procedures. Whether it adds useful information beyond an interview depends on the particular combination and should not be assumed from the fact that both were administered.

This publication’s Work Pattern Report illustrates the boundary. Its 100 items produce a low-stakes self-report across ten work-pattern continuums, with no norms, cutoffs, or selection score. It can help a person name questions about decisions, planning, feedback, conflict, collaboration, change, or learning. It has not been validated for hiring, promotion, pay, performance management, diagnosis, or career matching. Its findings therefore belong in private reflection or a coaching conversation, not an applicant ranking. A person considering a report should check its version, declared purpose, target group, scoring explanation, and validation evidence before assigning workplace importance to a result.

A responsible report interpretation also separates the construct from the recommendation. If a scale describes preference for order, that does not by itself imply the reader should choose a highly structured occupation, nor that they cannot work amid change. Context, skill, training, incentives, resources, and team design all affect behavior. A score may frame a useful follow-up: “Tell me about a time a plan changed and what you did next.” It should not replace that evidence when the decision is whether the competency is present.

Sources: Principles for the Validation and Use of Personnel Selection Procedures, Fifth Edition; Job Analysis; Personality Measures as Predictors of Job Performance: A Meta-Analytic Review

Open book with slider scales, arrows pointing toward profile cards, and a gauge beside a clipboard grid.
Open book with slider scales, arrows pointing toward profile cards, and a gauge beside a clipboard grid.

How would the same competency look in both methods?

Consider the defined competency of keeping a project on track when dependencies shift. In a structured interview, every candidate could receive the same job-related question about a comparable disruption, followed by planned probes: what changed, how did you reprioritize, whom did you inform, and what happened? A scoring guide could specify the evidence expected at different levels, such as identifying dependencies, communicating tradeoffs, revising milestones, and checking follow-through. This is an illustrative design, not a validated interview or a claim about any actual candidate. Its value comes from making the target visible and the evaluation criteria explicit.

A personality report might instead show how a respondent describes their preference for planning, flexibility, or closure, depending on what the instrument actually measures. The reader should verify those constructs in the manual rather than assuming a label means the same thing across reports. Such a result could suggest a follow-up: does the person plan in advance, improvise effectively, or switch between approaches depending on constraints? It cannot establish that a project stayed on schedule, because a preference is not the same thing as performance and self-report is not a record of the work outcome.

The two methods can be compared against the same decision without pretending they produce the same kind of evidence. If the intended decision is coaching, a report may help the person notice a pattern and choose an experiment: record how they respond the next three times priorities change, including what they did and what the situation required. An interview can surface examples and reasoning, but an interviewer’s impression remains an interpretation. If the decision is hiring for immediate planning capability, the interview should be designed around relevant behavior and may be supplemented by a work-like task when that task provides more direct evidence. The best method depends on what the job actually demands, not which instrument sounds more scientific.

This comparison also prevents two common mistakes. First, it prevents the report from being treated as a skills test merely because its narrative mentions planning. Second, it prevents the interview from being treated as objective simply because questions are standardized. The interview can still reward rehearsed stories, verbal fluency, or familiarity with the format unless the design and scoring focus on relevant evidence. A practical record should distinguish what was observed or reported, the criteria applied, and the inference drawn. For example: “described identifying the blocked dependency and notifying the team” is evidence; “therefore always plans well” is an overgeneralization.

Sources: Structured Interviews; Principles for the Validation and Use of Personnel Selection Procedures, Fifth Edition; Comprehensive Review and Meta-Analysis of the Validity of Interviews

What fairness and job relevance require

A consistent process can still be unfair or poorly matched. Questions may rely on experiences that some candidates had little opportunity to acquire, or a scoring guide may reward a communication style unrelated to the job. A self-report measure may create its own access, language, interpretation, and response concerns. Fairness is not settled by calling a process structured, validated, or personality-based. It requires attention to who can demonstrate the required behavior, what the procedure actually asks of them, and whether the result is used in the stated way.

U.S. Equal Employment Opportunity Commission guidance explains that employment practices with significant adverse impact may need to be job-related and consistent with business necessity, and discusses the need to consider less discriminatory alternatives in relevant contexts. The Uniform Guidelines on Employee Selection Procedures provide a federal framework for validity evidence and selection procedures. These are U.S.-specific sources, not a full legal analysis for every employer or jurisdiction. Their practical lesson for this article is restrained: selection decisions carry obligations beyond choosing the instrument with the clearest narrative, and documented job relevance matters. Employers should obtain qualified guidance for the applicable law and context.

For an interview, fairness review includes whether all applicants receive equivalent core questions, whether reasonable accommodations are available, whether scoring criteria are tied to actual work, and whether raters are trained to apply them. For a personality measure, review includes whether the instrument is appropriate for the population and language, whether evidence matches the intended purpose, what happens to results, and whether the process invites unsupported inferences about protected or irrelevant characteristics. Similar administration does not guarantee equitable outcomes; outcome patterns and accessibility should be monitored where appropriate.

Readers using reports for self-reflection face a lower-stakes but related interpretive obligation: do not confuse difference with deficiency. A less preferred style is not automatically poor performance. A person who prefers to improvise might meet a deadline through reminders and team coordination; a strong planner might still struggle when requirements are unstable. The question is whether the behavior meets the actual requirement under the actual conditions. In coaching, pair a report prompt with concrete examples and contextual questions. In selection, require stronger evidence and safeguards, because the decision affects another person’s access to work.

Sources: Section 15: Race and Color Discrimination; Uniform Guidelines on Employee Selection Procedures, 29 CFR Part 1607

So which should you use for a specific competency?

For a specific job competency, begin with a structured interview designed from job analysis when the competency can be elicited through past behavior or a relevant scenario. It is generally the more direct of these two methods for asking what a person did or would do in a defined work situation. Use a personality report to understand tendencies or generate questions, especially for reflection and development. Treat it as selection evidence only when the particular instrument, score interpretation, population, and use have suitable supporting evidence. If the competency is better shown by doing the task, consider whether a work sample is more direct than either option.

A simple decision sequence helps. First, state the competency in observable terms and the decision it informs. Second, identify evidence that would count: an action, result, knowledge demonstration, or recurring tendency. Third, inspect each method’s content and scoring process. Fourth, check that the evidence matches the intended use and relevant population. Fifth, record uncertainty and avoid turning one score or interview answer into a broad statement about a person. For a development question, the threshold may be a useful, checkable hypothesis. For selection, the evidence and fairness requirements are more demanding.

The strongest case for a report is that a well-supported, job-relevant personality measure can contribute information that a single interview may not capture. The strongest case against relying on it for this question is that a tendency score does not directly demonstrate a named capability, and broad research cannot validate a particular report for a particular decision. The exception matters: if the target is a relevant personality tendency rather than a skill, the instrument is suitable for the population, and the intended use has evidence, it may be pertinent. Even then, do not infer more than the evidence warrants.

The verdict changes the opening question from “Which format is better?” to “What kind of evidence would answer my competency question?” For a planning competency, ask for comparable examples and score the actions against clear standards; use a report, if appropriate, to prompt a separate conversation about preferred patterns and conditions. The Work Pattern Report at /assessment can help an individual organize that low-stakes reflection across planning and other work continuums. It supplies no hiring score or job recommendation. A useful next step is to write one observable behavior you want to understand, then gather evidence of that behavior in context.

Sources: Structured Interviews; Principles for the Validation and Use of Personnel Selection Procedures, Fifth Edition; Personality Measures as Predictors of Job Performance: A Meta-Analytic Review

Questions readers ask

Can a personality test predict whether someone will perform well at work?

Some personality measures have been associated with job performance in studied groups, and job analysis can help identify relevant measures. That evidence does not show that every report predicts every role or proves an individual’s specific competency. Check the instrument, population, criterion, and intended use.

Sources and notes

  1. Structured Interviews

    OPM defines the method, common questions, and shared rating standards for job-related competency interviews.

  2. Job Analysis

    OPM explains how task and competency analysis anchors selection decisions and establishes job relevance.

  3. Principles for the Validation and Use of Personnel Selection Procedures, Fifth Edition

    Professional selection principles describe matching procedure content and validity evidence to work requirements and intended use.

  4. Comprehensive Review and Meta-Analysis of the Validity of Interviews

    The meta-analysis reports that interview validity varied by content, structure, format, and criterion in the reviewed literature.

  5. Identification and Meta-Analytic Assessment of Psychological Constructs Measured in Employment Interviews

    The study reviews interview constructs and proposes that construct content partly explains validity differences.

  6. Personality Measures as Predictors of Job Performance: A Meta-Analytic Review

    The review finds personality-performance associations varied and were higher in studies using job analysis to select measures.

  7. The Big Five Personality Dimensions and Job Performance: A Meta-Analysis

    The abstract raises measurement-composition concerns in prior Big Five syntheses and supports caution about broad generalization.

  8. Section 15: Race and Color Discrimination

    EEOC guidance discusses job-relatedness, validation, adverse impact, and alternatives in U.S. employment selection contexts.

  9. Uniform Guidelines on Employee Selection Procedures, 29 CFR Part 1607

    The federal guidelines describe a validity and equal employment opportunity framework for employee selection procedures.

Apply it to your work

Turn a work-style question into specific observations

From this guide: A report can suggest what to notice, but your decision still depends on how a tendency appears in the work situations that matter to you.

If you are weighing a career or collaboration decision, the remaining question is how your planning, decision, feedback, and change preferences show up in actual situations. The Work Pattern Report offers a low-stakes self-report across those continuums so you can turn a broad question into observations to examine. Use it for reflection and discussion, not as a job match or employment score.