Yes, a personality report can support a coaching goal without being the measure of improvement. Its useful role is limited: it may give a person language for a recurring tendency, prompt a conversation about a goal, or suggest questions to test against real situations. It does not, by itself, establish that the tendency caused a problem, that coaching changed it, or that an important outcome improved. Keep the report as a starting hypothesis. Define the coaching goal in observable terms, record a baseline, choose one or more measures, and review the same measures later. If the report is retaken, treat the new score as one piece of information rather than a verdict on progress.
The claim under review: insight is not improvement
A person may read a report that describes a tendency toward planning, social interaction, caution, or novelty seeking and recognize a useful coaching question. For example: when a project changes suddenly, do I pause too long before acting? That question can help a coach and client choose what to examine. The report has supported the goal by giving the conversation a possible starting point.
The stronger claim is different: because the report identified the tendency, coaching will change it, and a later report will prove improvement. That conclusion does not follow. A score is an observation produced by a particular instrument, response process, scoring rule, and reference frame. It is not a direct recording of every behavior in every setting. The meaning of the score also depends on the instrument's intended use and the evidence supporting that interpretation.
The practical verdict is therefore narrow. Use a personality report to generate and refine questions. Do not use it as the sole outcome measure for a coaching engagement, especially when the goal concerns a concrete behavior such as preparing for difficult conversations, delegating work, or meeting deadlines. Those outcomes need their own evidence.
What a report can contribute to a coaching goal
A report can contribute in three modest ways. First, it can provide a shared vocabulary. Instead of arguing about whether someone is simply difficult, a coaching conversation can examine a more specific tendency and ask when it appears, when it does not, and what the situation demands. This keeps the discussion closer to observations than labels.
Second, it can help identify a question worth testing. A report that describes a high or low position on a scale does not tell a person what to do. It may, however, invite a comparison with recent examples: What happened before the meeting? What did I do next? What did another person observe? What would a more effective response look like? Agreement with a report is not proof of accuracy, and disagreement is not proof that the instrument failed. Both are prompts to inspect the evidence.
Third, it can help a client choose a focus that feels relevant. A coaching goal works better when the client can state why it matters and what would be different in daily life. The report may make a concern easier to name, but the client still owns the goal. A coach should not convert a scale into an obligation to become more outgoing, more agreeable, or more controlled. There is no universally good direction on a general personality scale.
This is a different role from screening or selection. A report used for self-reflection or coaching should not quietly become a diagnosis, a job judgment, or a claim about what someone can never learn. The purpose of the score interpretation must match the evidence for that purpose.
Why the report cannot measure the coaching outcome alone
A personality report and a coaching outcome answer different questions. The report may ask how a person generally describes a pattern across situations. An outcome measure asks whether a specified behavior, skill, relationship, or result changed over a defined period. The first can inform the second, but the two should not be treated as interchangeable.
Consider a goal to contribute earlier in project meetings. A report might suggest that the client is cautious in unfamiliar groups. That may be a useful hypothesis, but it leaves important questions open. Is the barrier uncertainty about the topic, the power difference in the room, lack of preparation, fatigue, or a fear of interrupting? What counts as improvement: speaking once, offering a recommendation, asking a clarifying question, or helping the group reach a decision? A broad scale cannot answer all of those questions.
A second problem is attribution. If the client speaks more often after coaching, several explanations may be possible: practice, a new manager, a change in meeting format, a clearer role, or ordinary variation. The change may be valuable without being caused by a shift in a personality score. The American Psychological Association's summary of research on personality-feedback interventions makes a similar caution: feedback is often combined with coaching, training, or other programs, so studies have difficulty attributing outcome improvements to the feedback itself.
A third problem is precision. Even a well-designed measure contains uncertainty. A small difference between two administrations may reflect normal variation, changes in context, or measurement error rather than a meaningful change in the underlying tendency. A report should explain its score type, comparison group, reliability information, and limits before a reader treats a change as decisive.

A worked example: from report language to a measurable goal
Suppose a report describes a tendency related to organization. The unhelpful coaching goal is: become more conscientious. It is broad, value-laden, and difficult to evaluate. It also encourages the person to chase a label instead of examining what happens in a real setting.
A narrower version might be: before the weekly project meeting, send a short agenda and identify the two decisions that need attention. The behavior is visible. The client and coach can agree on a time frame, such as the next four meetings, and record whether the preparation happened. They might also track whether the meeting ended with clear owners and next steps. Those measures do not prove that a personality trait changed. They show whether the chosen practice occurred and whether it was associated with a useful meeting result.
The report can still help. It may prompt questions about planning, follow-through, or how the client responds when priorities compete. The coach can test those interpretations against examples and let the client reject a story that does not fit. The goal then belongs to the situation and the behavior, not to a demand to move toward a preferred personality category.
The same logic works for a social goal. Replace become more extraverted with ask one substantive question in each stakeholder meeting and note whether the question improved understanding. Replace be less agreeable with state one considered disagreement respectfully when the evidence warrants it. Replace be less anxious with identify the physical or situational cue, use an agreed coping response, and record whether the person returned to the task. These are examples of goal design, not claims that a particular report predicts the behavior.
What to measure instead of a personality score
Start with the intended outcome. Coaching guidance from the International Coaching Federation recommends agreeing on goals and measures at the beginning, tracking actions and progress during the engagement, and considering later follow-up. That sequence is more informative than waiting until the end and trying to reconstruct what changed.
Choose a small set of measures that fit the goal. An action measure records whether the planned behavior happened. A quality measure records how well it met an agreed standard. An outcome measure records a consequence that matters to the client. For a delegation goal, these might be the number of tasks handed over with a clear brief, the recipient's understanding of the handoff, and whether work was completed without avoidable rework. No single measure captures the whole story.
Use the same definitions at the start and later. A simple record can include the date, situation, intended behavior, what happened, and what the client learned. For work goals, relevant feedback from a manager, peer, or direct report may add a perspective, but confidentiality and consent should be agreed in advance. A self-rating can be useful for reflection while still being only one source of evidence.
A later personality assessment may be worth discussing if the instrument is designed and documented for repeated use. Even then, interpret the result cautiously. A changed score may reflect genuine change, a changed frame of reference, response conditions, or ordinary score uncertainty. If the goal is behavior, behavior remains the primary evidence.

What the research does and does not show
Research does not require the conclusion that personality is fixed. A systematic review by Roberts and colleagues examined changes in personality-trait measures during interventions and found evidence of change across the studies it included. That finding supports the possibility that measured tendencies can change over time. It does not show that any particular commercial report will detect a client's improvement, or that a coaching conversation caused a later score difference.
More focused intervention research illustrates why the measurement question matters. A digital-coaching study examined self-reported and observer-reported changes at the domain, facet, and item levels. The reported patterns were not uniform across those levels, and observer-reported changes were generally small and not statistically significant in the abstract. A broad total or domain score can therefore hide which narrower pattern changed, if any.
A controlled field experiment on workplace coaching measured performance with self-ratings and supervisor ratings over three time points and examined how individual differences related to coaching effects. This is closer to an outcome design than simply repeating a personality questionnaire, because it included a comparison group and performance measures. It still answers a specific research question in a particular sample and setting, not a universal rule for every report or coaching goal.
The safest synthesis is not that personality reports are useless or that they prove transformation. They can be useful conversation tools when their intended use is clear and their interpretations remain tentative. Evidence for improvement should come from measures chosen for the goal, collected at sensible points, and interpreted with attention to alternative explanations.
The report-reading checklist
Before using a personality report in coaching, ask what the instrument measures and whether the report states its intended use. Find out whether the result is a raw score, a standardized score, a percentile, or a band, and identify the comparison group behind any norm-referenced interpretation. A percentile describes a position in a reference distribution; it is not a percentage of a trait and is not a grade.
Ask what evidence supports the interpretation you are being invited to make. Reliability concerns the consistency or precision of scores under specified conditions. Validity concerns whether evidence supports a proposed interpretation and use. A reliable score can still be used for the wrong purpose. If a publisher offers only an attractive narrative and no usable technical information, keep the conclusions modest.
Then translate the report into a testable coaching question. Write down the situation, the behavior you want to observe, the reason it matters, and what would count as progress. Select a baseline before practice begins. Decide when you will review the record and who, if anyone, may contribute feedback. Do not make a later score the only pass or fail decision.
Finally, protect the boundary around the report. Use it for reflection or development only when that is the agreed purpose. Do not infer a diagnosis, moral worth, fixed ability, or employment suitability from a general personality result. If the report is being used by an organization, ask who can see it, how it will be interpreted, and what decisions it may influence.
Questions readers ask
Can a personality report help me choose a coaching goal?
Yes. It can offer language for a possible tendency and prompt questions about situations where that tendency matters. Confirm the interpretation against observable examples, then write the goal as a behavior or outcome rather than as a demand to change a personality label.
Should I retake the personality assessment after coaching?
Only if the instrument is documented for repeated administration and the timing makes sense. Treat the new score as one piece of evidence. Compare score uncertainty and administration conditions, and keep the original behavior-based measures central to the review.
What is a better measure of coaching improvement?
Use measures tied to the goal: whether a behavior occurred, the quality of that behavior, and a relevant outcome. Set a baseline, use consistent definitions, review progress during the engagement, and add consented feedback from other observers when it is appropriate.
Does a changed personality score prove that coaching worked?
No. A score difference may reflect real change, measurement uncertainty, response conditions, or other events. It also cannot by itself show that coaching caused the difference. Stronger conclusions use goal-specific measures and consider alternative explanations.
Sources and notes
- Standards for Educational and Psychological Testing
The joint AERA, APA, and NCME standards frame validity, reliability, score interpretation, and test use as distinct professional concerns.
- Personality-feedback interventions have ambiguous effects on performance
The APA summary reports limited evidence for beneficial performance effects from personality feedback and difficulty attributing outcomes when feedback is combined with other interventions.
- Five Steps to Evaluate Your Coaching Practice
ICF guidance describes setting success criteria, choosing measures, selecting measurement points, communicating results, and reviewing the evaluation process.
- Measuring the ROI of Coaching: A Pragmatic Approach for Coaches
ICF guidance recommends agreeing on coaching goals and measures early, then recording specific behavioral changes and relevant impacts.
- A systematic review of personality trait change through intervention
The systematic review examined measured personality-trait change during interventions and supports the possibility of change without validating every report or coaching use.
- Personality change through a digital-coaching intervention
The study examined self- and observer-reported change at domain, facet, and item levels, illustrating why broad score changes can conceal uneven patterns.
- The Effects of Coachee Personality and Goal Orientation on Performance Improvement Following Coaching
This controlled field experiment measured coaching-related performance with self and supervisor ratings across three time points and a comparison group.
- APA Guidelines for Psychological Assessment and Evaluation
The APA guidelines emphasize judging validity evidence for the intended purpose and considering positive and negative consequences of score use.
Apply it to your work
Understand how you work before you choose what comes next.
From this guide: Carry this report-reading question into the work decision in front of you.
Build a private Work Pattern Report across ten workplace continuums, then compare the result with the demands of the role or environment in front of you.
