Skip to content

Criteria-First Evaluation Brief

Document — instantiates Identity-Safe Performance Context

Communicates the target capability, evidence standard, process, supports, limits, and review rights before evaluation.

When the rules of judgment are unstated, a participant has to infer them — and under identity threat, the cheapest inference is that some unspoken, identity-coded notion of "fit" is what really decides the outcome. Criteria-First Evaluation Brief removes that inference by putting the rules in writing and in the participant's hands before the attempt: what capability is being measured, what evidence will count, how the process runs, what support is allowed, where the hard limits are, and how to seek review. It is an artifact of transparency — a document that makes the standard legible up front. That is its defining move and its boundary: it publishes the standard to participants; it does not align the evaluators who apply it (that is Evaluator Exemplar Calibration), and it does not deliver the standard back as feedback afterward (that is the Process-Evidence Feedback Template).

Example

A software company replaces its old "we'll know a good engineer when we see one" screening with a work-sample exercise, and pairs it with a one-page brief sent to every candidate the day before. The brief states the target capability plainly: "We are assessing whether you can diagnose and fix a bug in an unfamiliar codebase and explain your reasoning — not speed, not memorized syntax, not familiarity with our stack." It lists the evidence standard: the two things reviewers must see (a correct-enough fix and a legible explanation), what is explicitly out of scope (coding style, which editor you use), and that reviewers score the artifact against a shared rubric, not their gut. It says you may use documentation and search, that you have ninety minutes, and that if you believe a prompt was unclear or unfair you can flag it through a named channel without it counting against you.

The setup precedes the exercise by design. The intended outcome is that a candidate who is a strong engineer but an anxious interviewee spends the ninety minutes on the bug rather than on decoding whether "culture fit" is the hidden test — because the brief has already told them it isn't.

How it works

The brief works by pre-committing the evaluation's terms and making the commitment visible:

  • Separate the target from the traditional. It states what the capability genuinely requires and, just as important, names common demands that are not part of it (accent, polish, insider vocabulary, improvisational confidence), so those cannot silently become the standard.
  • Make evidence concrete. It says what artifacts and observations will drive the decision and which standards are non-negotiable, so the participant does not have to guess whether identity-coded impressions will weigh in.
  • Bound the process. Timing, allowed supports, and limits are stated, converting ambient uncertainty into known constraints.
  • Publish the exits. Review rights and the channel to use are named, so recourse is knowable in advance rather than discovered in a crisis.

Tuning parameters

  • Specificity vs. gameability — how precisely the evidence standard is spelled out. More precision reduces identity-coded inference but makes a process easier to rehearse to; calibrate to how much a coached answer would actually invalidate the measure.
  • Scope of disclosed limits — how much the brief names what is not being measured. Naming more construct-irrelevant demands protects participants but lengthens the document and can invite line-item argument.
  • Reading load — one tight page vs. a full protocol. Shorter is more likely to be read; longer is more complete. The failure is a brief so exhaustive no one opens it.
  • Support disclosure — how explicitly permitted aids and accommodations are stated. Stating them widens genuine access but can shift where candidates spend effort.
  • Timing of release — far ahead vs. just before. Earlier release lowers anxiety and aids preparation; too early risks the brief being forgotten by attempt time.

When it helps, and when it misleads

Its strength is that it collapses the space in which identity-coded "fit" can hide, and it does so before performance, when it can still change how attention is spent. Making the criteria public is a direct application of procedural transparency — people extend more trust and spend less energy self-monitoring when they can see that the rules are stable and impersonally applied.[n1]

Its failure mode is cosmetic transparency: a brief that reads fairly while the actual decision still runs on undisclosed impressions. If the published standard and the evaluators' real behavior diverge, the document becomes evidence that the process is managed for appearance — worse than no brief at all. A subtler misuse is over-specification that games the measure: spelling out the rubric so exactly that a coached candidate can produce the target answer without the target capability. The guarding discipline is to bind the brief to what evaluators are actually held to — which is precisely why it needs a calibration mechanism downstream — and to audit whether decisions track the published evidence rather than something else.

How it implements the components

  • target_capability_boundary — the brief's opening job: it states what success requires and, explicitly, what is not part of the capability, drawing the line participants and evaluators both work inside.
  • transparent_evidence_standard — it publishes, before the attempt, what evidence determines the decision, which standards are fixed, and how review works, so participants need not infer whether identity will influence judgment.

It does NOT align evaluators to apply that standard consistently (evaluator_calibration_boundary) — that is Evaluator Exemplar Calibration; nor does it deliver the standard back as post-performance feedback (process_focused_feedback_channel), which is the Process-Evidence Feedback Template.

Editorial Notes

Form Classification

Form family: Communication, Facilitation & Learning

Rationale: Criteria-First Evaluation Brief operates as a designed message, facilitated interaction, ritual, or learning activity that changes shared understanding because it communicates the target capability, evidence standard, process, supports, limits, and review rights before evaluation.

Independent corroboration: The frozen evidence defines Criteria-First Evaluation Brief as 'Communicates the target capability, evidence standard, process, supports, limits, and review rights before evaluation', so its operative form is Communication, Facilitation & Learning.

Nearest alternative: Rule, Policy & Commitment — The pre-evaluation brief operates as a designed message that makes terms legible, while formal enforcement lies elsewhere.

Review outcome: Independent reviewer agreement; medium confidence.

Origin Attribution

Primary origin: Education & Pedagogy

Origin pattern: Cross-disciplinary synthesis

Present-day reach: Multi-domain

Rationale: Assessment pedagogy cohered advance disclosure of learning targets, evidence standards, allowed supports, and evaluation process to make judgment transparent.

Related originating lineages:

  • Law & Governance — Procedural fairness supplied advance notice of criteria, consistent process, reasons, and review rights.
  • Organizational & Management Science — Structured selection supplied predeclared capability targets, evidence standards, supports, and constraints.
  • Psychology — Identity-threat and procedural-justice research supplied the rationale for making hidden standards explicit before performance.

Review resolution: The brief is an educational assessment artifact synthesized with legal procedural fairness, structured selection, and identity-safety psychology.

Encyclopedia synthesis: The exact catalogued form synthesizes established practice rather than reproducing a single standard historical label.

Review outcome: Reconciled after independent review; high confidence.

Notes

[n1] Procedural justice — the finding that people judge fairness heavily by whether the process is consistent, transparent, and impartial, not only by whether they win. A published, stable evidence standard is the procedural-justice lever this document pulls.