Glass jar layered with stones and sand, flanked by small bowls of materials.

The problem with turning a score into an identity

A score summarizes evidence obtained under certain rules, sources, and conditions. When language turns it into a definition of who the person is, a time-bound measurement begins to govern expectations, opportunities, and new interpretations.

PRYSMAP4 min read

The transformation can happen in a single sentence. The result placed someone at level 2 becomes they are level 2. An observation about performance against a set of expectations begins to describe what kind of person they are. The capability, time frame, evidence, and intended use are lost.

A score needs a complete sentence

An isolated number appears self-sufficient. To interpret it, however, one must know at least:

  • which skill or manifestation was assessed;
  • the expectation used as a reference;
  • with which sources and evidence;
  • during which period and in what context;
  • the rules used to produce it;
  • for which decision it is considered relevant.

The Standards for Educational and Psychological Testing are explicit: validity refers to the interpretation of scores for proposed uses; no test permits valid interpretations for every purpose or situation (AERA, APA, and NCME, 2014 (opens in a new tab)). A score does not carry its meaning within itself. It acquires meaning through an argument connecting evidence, interpretation, and use.

That is why language should preserve the complete sentence even when it does not repeat it in full every time: the evidence available during this period supports a given conclusion for this decision, within these limits. Compression may be necessary in a dashboard; the interpretation should not disappear from the conversation.

The label reshapes future evidence

Once someone is labeled a low performer, high potential, or not strategic enough, new episodes are interpreted within an existing narrative. Behaviors consistent with the label are easily remembered. Those that contradict it may be treated as exceptions. Opportunities also change: the person may stop receiving assignments in which they could produce different evidence.

No bad intent is required. Teams need to simplify information and coordinate decisions. A stable category makes that task easier. The cost appears when administrative convenience makes the assessment circular: the score limits experiences, and the lack of experience confirms the score.

The person assessed may also begin to interpret their work through the category. If they receive a number without concrete manifestations, they do not know what to preserve, change, or question. The result functions as a verdict. If they receive evidence, context, and limits, they can challenge the interpretation and take part in a development decision.

Hiding the result does not solve the problem

Eliminating scores or replacing them with vague language does not guarantee a fairer conversation. Informal categories may be less traceable than an explicit scale. Excellent, role model, or not yet mature can also become identities, with the added problem that their rules are usually harder to reconstruct.

The goal is to limit what the measure can claim. This requires both design and disciplined use:

  • name the capability and expectation, not some overall quality of the person;
  • separate insufficient evidence from low performance;
  • show the manifestations that support the conclusion;
  • keep the date or observation window visible;
  • restrict decisions for which the result was not validated;
  • allow challenge, new evidence, and reassessment.

The APA guidelines on psychological assessment also address the responsibilities of those who produce, use, interpret, and provide feedback on results. That breadth is a reminder that quality does not end with the instrument; it continues through the way the conclusion is communicated and used (APA, 2020 (opens in a new tab)).

Black suitcase on a conveyor belt with a white tag on its handle.

From identity to a revisable state

Result A conclusion based on particular sources, rules, and conditions.

Interpretation The scope that the evidence allows it to support.

There is a substantive difference between saying they lack coordination capability and we do not yet have sufficient evidence of autonomous coordination in projects with multiple dependencies. The second formulation is not kinder as a matter of courtesy. It is more precise: it identifies the manifestation, the condition, and the limit.

It also changes the next step. An identity invites people to accept or reject the judgment. A revisable state makes it possible to ask which experience would produce new evidence, what support would help, and when reassessment would make sense. It does not promise that the conclusion will change; it opens the possibility of testing it.

This does not prevent difficult decisions. An organization may conclude that the current evidence does not meet the expectation of a role or that someone is not ready for an immediate transition. Precision does not remove the consequence. It prevents the consequence from extending beyond what is known.

A score serves a valuable function when it compresses information without erasing its origin. It stops doing so when it replaces the person’s name or spills into decisions it was never designed to support. The challenge is not to make the number say more. It is to preserve the discipline of remembering what it can tell us—and what remains beyond its scope.