Skip to content

Observation Warrant Ladder

Method — instantiates Appearance vs. Reality Distinction Audit

Defines levels of claim strength from raw report through corroborated measurement to robustly inferred reality claim.

Observation Warrant Ladder is an ordinal scale of claim strength. Its rungs run from a raw single report, to a corroborated observation, to a measured indication, to a robustly inferred reality claim — and the point of the ladder is not just to name the rungs but to specify what evidence promotes a claim from one rung to the next. Its defining move is that it ranks strength, not kind: it says nothing about whether a claim is an experience or a measurement, only about how far the current evidence licenses it up the confidence gradient. It is a shared yardstick the rest of the audit reaches for whenever it needs to say how strong a claim is, and it exists precisely to prevent the two opposite errors — overclaiming (jumping to the top rung) and permanent skepticism (refusing to ever leave the bottom).

Example

A team hunting exoplanets sees a small periodic dip in a star's brightness. The ladder keeps them honest about what they can say. Rung 1 — raw report: one telescope logs a single dip; this is a candidate, nothing more. Rung 2 — corroborated observation: the dip recurs at the same interval across several transits and survives checks for instrument artifacts. Rung 3 — measured indication: an independent method — a radial-velocity wobble of the star — yields a consistent mass, so two different measurements now point the same way. Rung 4 — robustly inferred reality: the convergent evidence clears the field's discovery bar and the planet is announced as real.

At each rung the wording the team is allowed changes: "a candidate signal," then "a strong candidate," then "consistent with a planet of roughly this mass," then "a confirmed planet." The ladder makes the promotion criteria explicit, so nobody upgrades the language until the evidence for the next rung is actually in hand.

How it works

  • Fix the rungs. Name an ordered set of warrant levels for the domain, each with a one-line description of what a claim at that level is and is not entitled to say.
  • Bind each rung to a source and a test. Tie each level to the kind of evidence that characterizes it and, crucially, to the specific corroboration required to climb out of it.
  • Grade the claim to its current rung. Place the claim at the highest level its present evidence actually supports — no higher.
  • Cap the wording at the rung. The rung sets a ceiling on the claim's verb; a rung-2 claim may not be phrased as a rung-4 certainty.
  • Re-grade on new evidence. Promotion requires meeting the next rung's corroboration bar; a claim can also be demoted when a corroboration fails to replicate.

Tuning parameters

  • Number of rungs — a coarse three-level ladder versus a fine six-level one; more rungs express finer distinctions but slow grading and invite quibbling over neighbors.
  • Promotion bar — how much corroboration each climb demands; a high bar resists overclaiming but risks freezing claims below their due.
  • Top-rung calibration — where "robustly inferred reality" is set for the domain; a courtroom, a physics lab, and a product team draw it very differently.
  • Symmetry — whether the ladder allows demotion as readily as promotion; an up-only ladder ratchets toward overconfidence.
  • Bridge visibility — whether each promotion must name the inference it rests on or may climb silently.

When it helps, and when it misleads

Its strength is that it gives a team a shared, ordered vocabulary for disagreement: two people who both use the ladder can locate exactly which rung they dispute, instead of trading "it's real" against "you can't know that." It disciplines overclaiming and reflexive doubt in the same stroke.

Its failure mode is at the extremes. Set the top rung unreachably high and the ladder manufactures permanent skepticism, treating nothing as ever established; set the promotion bar too low and it launders weak claims into confident ones. Even a rigorous bar can mislead — particle physics fixed a demanding five-sigma threshold for announcing a discovery,[n1] yet a five-sigma bump can still be a fluctuation or a look-elsewhere artifact, so a high rung means strongly warranted, never certain. The classic misuse is reading the top rung as proof and closing the case. The guarding discipline is to treat every rung as a statement of current warrant that new evidence can move — up or down — rather than a permanent grade.

How it implements the components

Observation Warrant Ladder realizes the audit's strength-grading machinery — the components that answer "how far does the evidence reach?" rather than "what kind of claim is this?":

  • warrant_boundary_statement — each rung is a warrant boundary; grading a claim to a rung states what the evidence supports and where it stops.
  • measurement_corroboration_plan — the promotion criteria specify the corroboration a claim must gather to climb, which is a corroboration plan in ladder form.
  • evidential_source_trace — each rung is bound to the source type that characterizes it, so climbing forces attention to where the evidence comes from.

It does not sort claims by type — the report/measurement/inference appearance_reality_classifier is Appearance/Reality Audit Checklist's job — and it does not turn an abstract claim into the tests that would give it content; that downward rewrite and its experience_condition_set belong to Sense-Condition Rewrite Template.

Editorial Notes

Form Classification

Form family: Rule, Policy & Commitment

Rationale: Observation Warrant Ladder operates as a standing rule, threshold, contractual commitment, or policy constraint governing future conduct because it defines levels of claim strength from raw report through corroborated measurement to robustly inferred reality claim.

Independent corroboration: The frozen evidence defines Observation Warrant Ladder as 'Defines levels of claim strength from raw report through corroborated measurement to robustly inferred reality claim', so its operative form is Rule, Policy & Commitment.

Nearest alternative: Representation, Specification & Plan — Observation Warrant Ladder includes features of a static representation, map, specification, schema, or prospective plan that externalizes information, but its defining operation is a standing rule, threshold, contractual commitment, or policy constraint governing future conduct.

Review outcome: Independent reviewer agreement; medium confidence.

Origin Attribution

Primary origin: Philosophy

Origin pattern: Cross-disciplinary synthesis

Present-day reach: Universal

Rationale: Epistemology developed graded warrant from appearance and report through corroboration to justified inference about reality.

Related originating lineages:

Review resolution: Both independent reviews agree on primary origin philosophy; reconciliation resolves reported_ambiguity, alternate_origin_disagreement. Formative alternate lineages retained: statistics_experimental_design, history_historiography. The broader reach of later applications is kept separate as domain_reach=universal; origin_mode=cross_disciplinary_synthesis describes the historical relationship among lineages. Confidence is conservatively reconciled to medium, and encyclopedia_synthesis=true preserves the reviewers' boundary judgment.

Attribution caveat: Different sciences use different ladders; the unified artifact is a synthesis of a philosophical structure with professional thresholds.

Encyclopedia synthesis: The exact catalogued form synthesizes established practice rather than reproducing a single standard historical label.

Review outcome: Reconciled after independent review; medium confidence.

Notes

[n1] The convention in particle physics that a new-particle claim is announced only when the signal reaches five standard deviations above background — roughly a one-in-3.5-million chance of arising from fluctuation alone. It is the field's institutional top rung, deliberately conservative because the cost of a false discovery claim is high.