Yes. You can ask who developed the assessment, what it measures, which norm group supports its percentiles or bands, when that norm data was collected, and what evidence supports the report's interpretations. A transparent provider should be able to describe the instrument and point to a manual, technical report, or research record. The norm group answers a comparison question: compared with whom? The sources and validity evidence answer a different question: why should this score mean what the report says it means? If those details are missing, treat the report as a prompt for reflection rather than a precise statement about you.
Start with the decision you need the report to support
Two reports can both look polished and still deserve different levels of trust. One may be a carefully documented questionnaire intended for self-reflection. Another may present a percentile and a confident paragraph without saying how the comparison group was formed. The important first question is not whether the report feels accurate. It is what you plan to do with it.
If you are using a report to notice themes in your own answers, missing technical detail is a reason for caution, not necessarily a reason to discard every observation. If a report will influence hiring, promotion, admission, coaching goals, or a consequential personal decision, the bar is higher. You need evidence for that interpretation and use, not just a plausible description.
Ask for the sources and norm group because they let you inspect the bridge between a response and a conclusion. A source may describe the construct or the questionnaire. A norm group may explain a percentile. Neither, by itself, proves that a report's narrative is accurate for every individual. That distinction keeps a request for transparency from becoming a demand for certainty.
What should ‘the sources’ include?
The word sources can refer to several different things, so ask the provider to name them. First, identify the instrument: its title, version, author or developer, constructs, number and type of items, response format, and scoring method. A report cannot be evaluated well if the underlying measure is unnamed or changes without notice.
Second, ask for the technical documentation. A manual or technical report should explain how items were developed, how scores are calculated, what population was studied, and which uses are recommended. It should report evidence about reliability or precision and about the validity of the interpretations being offered. Reliability concerns how much a score is affected by specified sources of variation. It is not the same as proof that the score measures what the report claims.
Third, ask for the research behind the particular language in your report. A study supporting a broad trait description does not automatically support a prediction about job performance, relationship behavior, or future choices. The testing standards describe validity as evidence supporting an interpretation of scores for a proposed use, rather than as a permanent property stamped onto a test. In practical terms, ask: ‘Which evidence supports this conclusion and this use?’
Finally, ask whether the report is generated from fixed scoring rules, a local interpretation system, or a mixture of both. If an automated report turns several scales into a recommendation or a profile story, the provider should explain the rules and evidence for that combination. A named algorithm is not enough; the interpretation still needs support.
What a norm group actually tells you
A norm group is the reference population used to interpret a score. In a norm-referenced report, your result is located relative to that group. The group might be defined by age, language, country, occupation, education, or another characteristic, but the label should be specific enough for you to judge whether the comparison is relevant.
‘Compared with a large sample’ is not a sufficient description. You should be able to learn who was invited, who participated, when the data were collected, how the sample was recruited, and whether the provider used weighting or other adjustments. The standards also call for descriptive statistics and information about the precision of the norms. These details matter because a convenient online sample may not represent the people with whom you are being compared.
The fit is a judgment, not a moral verdict. A norm group that does not resemble you does not make your answers meaningless. It does mean that a percentile or low-average-high label may not carry the same meaning the report suggests. For example, a comparison group drawn from one language population may not be an appropriate reference for someone who completed a translated version unless the provider has evidence for the adapted form and its interpretation.
Also ask how old the norms are. People and testing contexts change, and an older reference distribution may need evidence that it remains suitable. A publication date alone does not settle the question. What matters is whether the provider explains the collection period and why those norms remain appropriate for the current report.
A practical request you can send
You do not need specialist language to ask for useful documentation. You can write: ‘Please tell me the exact instrument and version used, what it is intended to measure, and where I can read its technical documentation. What norm group was used for my percentile or band? Please include the population definition, recruitment method, data-collection dates, sample size, weighting or exclusions, and the score-conversion method.’
Add: ‘Which evidence supports the interpretation in my report, and is that evidence for self-reflection, coaching, development, selection, or another purpose? What reliability or measurement-error information applies to the score? If the report combines scales into a profile or recommendation, where are those rules and their supporting studies described?’
This request separates facts from interpretation. The first part asks what happened in scoring. The second asks whether the resulting meaning is defensible for your intended use. If the provider cannot disclose a full manual because of security or licensing, it can still provide a technical summary, references, and a clear explanation of the relevant norm and interpretation. Privacy protection is not a reason to hide the method entirely.
Save the answer with the report. Versions change. A report that identifies the instrument, norm reference, date, and interpretation rules is easier to revisit than a page of general statements whose source cannot be reconstructed.
How to read the answer when the provider gives numbers
Numbers can create an impression of precision. Read what each number is doing before deciding what it means. A raw score is a result on the instrument's original scoring scale. A standardized score transforms that result according to a stated reference system. A percentile describes relative standing within a comparison distribution. A band groups scores into categories. These are different outputs, and a band should not be treated as a diagnosis or a natural boundary between kinds of people.
Then check whether the comparison group belongs to the score you are reading. A provider may show one overall norm group while making claims about a subgroup, or it may display a percentile without naming the distribution behind it. Ask which norm applies to each scale, facet, or composite. If several forms, languages, or scoring versions are treated as interchangeable, ask for the evidence supporting that equivalence.
Look for uncertainty as well. Measurement error is the expected variation in an observed score that is not part of the underlying tendency the test aims to measure. A reliability coefficient summarizes one kind of error under particular conditions; it does not tell you that every individual score is exact. A report that places someone near a category boundary should explain whether small score changes could alter the label.
A useful report makes these distinctions visible. It does not use a precise-looking percentile to conceal an unclear norm group or turn a broad tendency into a prediction about a particular event.

Sources, norms, and validity answer different questions
It helps to keep three questions apart. The norm question is: relative to which population is this score interpreted? The measurement question is: how consistent or precise is the score under relevant conditions? The validity question is: what evidence supports this interpretation for this purpose? A strong answer to one question does not settle the others.
For instance, a report could have a clearly described norm group but weak evidence for using its score to select employees. It could have evidence that responses are consistent while offering little support for a dramatic story about relationships. It could cite a respected study about the instrument while using a different language version, scoring rule, or automated composite that the study did not examine.
The European Federation of Psychologists’ Associations’ test-review model treats description, reliability or precision, and several forms of validity evidence as parts of a broader review. That structure is useful for readers because it discourages a single-number verdict. It also points toward independent review, not only a publisher’s summary.
When you ask for sources, therefore, do not ask only ‘Is this test validated?’ Ask which score interpretation was studied, in which population, under which conditions, and for what decision. That wording is more demanding, but it produces a more honest answer.
What to do with a partial or evasive answer
A short answer is not automatically proof that an assessment is poor. Some providers have legitimate limits around copyrighted items, confidential individual data, or security-sensitive details. The key issue is whether they can still explain the instrument, its intended use, its norm reference, and the evidence behind the interpretations without asking you to trust a label.
Treat these responses as warning signs: the provider will not name the instrument or version; it describes the norm group only as ‘our users’; it gives a percentile without a comparison population; it cites popularity instead of technical evidence; or it claims that a score reveals fixed ability, diagnosis, or inevitable behavior. A refusal to distinguish self-reflection from selection is especially important when another person or organization will act on the report.
Match your response to the stakes. For private reflection, you can keep only the observations that help you ask better questions about your behavior, and hold the conclusions lightly. For coaching or development, agree on observable goals and use the report as one input rather than a verdict. For work or selection, ask what job-related evidence and fairness review support the use, and do not accept a general personality report as a substitute for a broader decision process.
If the documentation remains unavailable, the responsible conclusion is limited: you may know what the report says, but not how well its scoring and interpretation are supported. That is a reason to reduce the weight you give it, not a reason to invent missing evidence.
The report-reading checklist to keep
Before relying on a personality report, check six items: the exact instrument and version; the construct it says it measures; the intended use; the norm group and collection dates; the scoring and uncertainty information; and the research supporting the interpretation you plan to make. Add a seventh when relevant: evidence for the language, accommodation, subgroup, or automated report format you received.
A simple decision rule follows. If the instrument and norm group are clear, the evidence matches the interpretation, and the uncertainty is stated, you have a basis for cautious use. If one part is missing, narrow the claim. If several parts are missing, use the report only as a prompt for questions and do not let it carry a high-stakes decision.
So, can you ask for the sources and norm group behind your personality report? Yes, and the request is part of reading the report responsibly. You are not asking the provider to predict your life or reveal private test items. You are asking what was measured, compared, and supported. Start with the checklist above, record the response, and continue through the live topics library for the next question your report raises.
Questions readers ask
What if a personality report has no named norm group?
Ask the provider to identify the comparison population, collection period, sampling method, and score-conversion method. Until that information is available, do not treat a percentile or low-average-high label as a well-defined comparison. You can still reflect on the questions and your answers, but keep the report's broader conclusions provisional.
Sources and notes
- Standards for Educational and Psychological Testing
Supports the requirements for describing norm populations, sampling, dates, weighting, precision, intended use, validity evidence, scoring documentation, and cautions about misuse.
- EFPA Test Review Model, Version 2025
Supports separating test description, reliability or precision, validity evidence, intended context, and independent review when evaluating an assessment.
- The Standards for Educational and Psychological Testing
Confirms the collaborative professional authority and open-access location of the AERA, APA, and NCME testing standards.
- APA Dictionary of Psychology: norm-referenced test
Supports the plain-language definition of a norm-referenced interpretation as comparison with a specified reference group.
Apply it to your work
Understand how you work before you choose what comes next.
From this guide: Carry this report-reading question into the work decision in front of you.
Build a private Work Pattern Report across ten workplace continuums, then compare the result with the demands of the role or environment in front of you.
