Task-Based Wayfinding Test¶
Test / assessment — instantiates Predictive-Cue Wayfinding Design
A facilitated study in which representative agents attempt realistic tasks and are observed choosing routes from local cues alone, measuring whether honest navigation actually succeeds for real intents.
The Task-Based Wayfinding Test answers a question only real users can: when someone who actually wants X arrives, do the local cues get them there? Representative agents are given realistic tasks and observed as they choose among paths using nothing but the cues visible at each branch — where they hesitate, which label they pick, where they go wrong, whether they arrive. The defining idea is cooperative, individual-level observation of genuine intent: it is neither an adversarial hunt for traps nor an aggregate count of clicks, but a close watch of a few real people reasoning their way through the space. Its distinctive yield is the why behind a cue's success or failure, drawn from watching the choice happen.
Example¶
A public library redesigns its website and wants to know whether patrons can find things before it launches. It recruits eight people who resemble its real users — a retiree, a parent, two students, a job-seeker — and gives each a set of realistic tasks: "renew a book that's due tomorrow," "find a downloadable audiobook," "book a study room for Saturday," "get help citing a source." Each patron works from the live cue surface, thinking aloud, while a facilitator watches silently and notes every branch: which label they choose, where they pause, where they double back.
Patterns emerge that no aggregate could name. Six of eight go to "My Account" to renew and only then find it — the renew action was correct but the cue "My Account" under-promised it. Everyone hunting the audiobook clicks "eBooks," bounces, and re-tries "Digital" — the two labels compete without distinguishing their destinations. And the study-room task fails for five patrons because the only cue, "Facilities," reads to them as building maintenance. The test does not just show that these cues fail; by watching real intent meet real cues, it shows why, which is exactly what the redesign needs.
How it works¶
The distinguishing mechanics are task realism, real intent, and close observation:
- Real tasks, real intents. Participants pursue goals they might actually have, so the cues are judged against genuine intent rather than a designer's imagined one.
- Local cues only. The agent navigates from what is visible at each branch — no coaching — which isolates whether the cues themselves carry the scent.
- Observed at each decision point. The facilitator records the choice, the hesitation, and the recovery at every inventoried branch, capturing behavior a metric would flatten.
- Explanation, not just outcome. Because the reasoning is visible (often via think-aloud), the test yields the cause of a wrong turn, not merely its occurrence.
Tuning parameters¶
- Task realism and selection — how closely the tasks mirror real goals, and which goals are covered. Realistic tasks predict field behavior; contrived ones test the wrong thing.
- Participant representativeness — how well the sample matches real agents (novices, non-native readers, the intended audience). An insider sample passes cues that outsiders would fail on.
- Facilitation style — silent observation versus think-aloud versus prompting. Think-aloud surfaces reasoning but can alter it; prompting contaminates the very choice under test.
- Decision points probed — which branches the tasks route through; a test only measures the cues its tasks happen to exercise.
- Success criteria — what counts as arrival (found it, found it unaided, found it quickly). The bar defines what "the cues work" means.
When it helps, and when it misleads¶
The test is the mechanism that explains why cues fail — the reasoning behind a wrong turn — which is exactly what aggregate telemetry cannot supply and what a redesign most needs. It is the wayfinding form of tree testing, the information-architecture method of checking whether people can find things from labels alone, before visual design muddies the result.[n1]
Its failure modes are the classic ones of small-sample observation. A handful of participants can mislead if they are not representative, and an insider sample will glide past the jargon that stops real novices. Artificial lab tasks can pass cues that fail under genuine, distracted intent, and a facilitator who nudges — "did you see the menu at the top?" — quietly repairs the very cue under test, manufacturing a success the field will not reproduce. The guarding discipline is realistic tasks, a representative sample, and silent facilitation, with success defined as unaided arrival so the cues, not the moderator, do the guiding.
How it implements the components¶
decision_point_inventory— the test routes tasks through, and records behavior at, the enumerated branches where agents must choose.agent_intent_context— it grounds every judgment in a real task and intent, so cues are evaluated against what agents actually want.wayfinding_feedback_signal— each observed session is a rich behavioral record of choices, hesitations, and recoveries fed back into the design.
It observes cooperatively and per-agent but does not adversarially hunt for cues engineered to deceive (misleading_scent_guardrail, false_scent_exception_log) — that offense-oriented search is Misleading-Cue Red Team; nor does it aggregate population telemetry (scent_decay_monitor), which is Scent Clickthrough Trace Dashboard.
Related¶
- Instantiates: Predictive-Cue Wayfinding Design — supplies the observed, task-driven measure of whether real agents can navigate from local cues.
- Sibling mechanisms: Breadcrumb and Landmark Trail · Cue-Destination Alignment Matrix · Destination Preview Card · Link-Label Scent Audit · Misleading-Cue Red Team · Progressive Disclosure Preview · Route Recovery Pattern · Scent Clickthrough Trace Dashboard
Editorial Notes¶
Form Classification¶
Form family: Experiment, Test & Rehearsal
Rationale: Task-Based Wayfinding Test operates as an active test, trial, simulation, drill, or rehearsal that generates evidence through a deliberate attempt or perturbation because it a facilitated study in which representative agents attempt realistic tasks and are observed choosing routes from local cues alone, measuring whether honest navigation actually succeeds for real intents.
Independent corroboration: The frozen evidence defines Task-Based Wayfinding Test as 'A facilitated study in which representative agents attempt realistic tasks and are observed choosing routes from local cues alone, measuring whether honest navigation actually succeeds for real intents', so its operative form is Experiment, Test & Rehearsal.
Nearest alternative: Assessment, Review & Assurance — Task-Based Wayfinding Test includes features of a bounded evaluation of existing evidence or work that produces a finding or disposition, but its defining operation is an active test, trial, simulation, drill, or rehearsal that generates evidence through a deliberate attempt or perturbation.
Review outcome: Independent reviewer agreement; medium confidence.
Origin Attribution¶
Primary origin: Human-Computer Interaction
Origin pattern: Cross-disciplinary synthesis
Present-day reach: Universal
Rationale: Task based wayfinding test derives most directly from human-computer interaction's usability, wayfinding, and information-design tradition; its defining operation is to a facilitated study in which representative agents attempt realistic tasks and are observed choosing routes from local cues alone, measuring whether honest navigation actually succeeds for real intents.
Related originating lineages:
- Architecture & Urban Planning — Planning's spatial allocation, accessibility, and territorial-design tradition provides a formative adjacent lineage for the same task based wayfinding test operation.
- Computer Science & Software Engineering — Computer science and software-engineering practice supplies a parallel or contributing lineage for the mechanism's defining operation: a facilitated study in which representative agents attempt realistic tasks and are observed choosing routes from local cues alone, measuring whether honest navigation actually….
- Psychology — Experimental, clinical, and behavioral psychology supplies a parallel or contributing lineage for the mechanism's defining operation: a facilitated study in which representative agents attempt realistic tasks and are observed choosing routes from local cues alone, measuring whether honest navigation actually….
Review resolution: Both blind reviewers independently select human_computer_interaction as the primary historical origin for the concrete operation—A facilitated study in which representative agents attempt realistic tasks and are observed choosing routes from local cues alone, measuring whether honest navigation actually succeeds for real intents. The queued differences concern alternate origin disagreement, origin mode disagreement, domain reach disagreement, encyclopedia synthesis disagreement, not the primary lineage. I retain every alternate that either reviewer explains, without a numeric cap, and choose origin_mode=cross_disciplinary_synthesis because the reviewers' combined evidence identifies material construction from multiple disciplines. domain_reach=universal records later portability rather than multiplying historical origins; confidence=high is the conservative shared evidentiary level, and encyclopedia_synthesis=true preserves either reviewer's affirmative synthesis finding.
Encyclopedia synthesis: The exact catalogued form synthesizes established practice rather than reproducing a single standard historical label.
Review outcome: Reconciled after independent review; high confidence.
Notes¶
[n1] Tree testing (also "reverse card sorting") is an information-architecture method in which participants are asked to locate items using only a site's text hierarchy, with no visual design or search, to measure whether the labels alone make destinations findable. The task-based test generalizes this to any wayfinding space and any cue type. ↩