Visual-Weight Mockup Comparison¶
Test / assessment — instantiates Proportion / Scale Calibration
Compares alternative size distributions while controlling color, position, weight, and content as far as practical.
A Visual-Weight Mockup Comparison pits alternative size distributions of the same content against each other — holding colour, position, weight, and content as constant as practical — to find which distribution actually delivers the intended reading order, balance, or emphasis. Its defining idea is that it isolates size as the variable and measures its perceptual effect on hierarchy through controlled comparison, rather than testing physical fit or an apparent-scale illusion. It is the sibling that treats "which size relation communicates best?" as an empirical question with matched alternatives.
Example¶
A team designing a portfolio dashboard debates which element should dominate — the total-balance figure, the day's change, or the allocation chart. Instead of arguing taste, they build three matched mockups that differ only in size distribution: identical colours, identical positions, identical data, only the sizes reallocated. Then they run a quick reading-order and comprehension check — which number does the eye reach first, can people find the day's change in under three seconds, does anyone miss the allocation chart entirely? Mockup B (balance largest, change second) wins on comprehension. Mockup C's giant allocation chart drew the eye first but people misread it as the headline number, mistaking prominence for meaning. Size assignments get locked from evidence, not from the loudest opinion in the room.
How it works¶
- Create matched alternatives that vary only in size distribution, holding other cues fixed.
- Randomize presentation order to blunt position and order effects.
- Gather reading-order, comprehension, preference, and salience data, since size is a strong preattentive cue and the eye goes to the largest thing first.[n1]
- Inspect subgroup differences to confirm the winning distribution works across audiences, not just on average.
Tuning parameters¶
- Control tightness — how many co-varying cues (colour, position, weight) you hold fixed. Tighter isolation attributes the effect to size cleanly but reads less like the real product.
- Evidence type — preference versus task performance (reading order, comprehension, find-time). Preference alone rewards spectacle.
- Alternative spread — how different the competing distributions are; too similar yields no signal, too wild tests nothing usable.
- Subgroup inspection — whether you verify the winning distribution across audiences rather than trusting the average.
When it helps, and when it misleads¶
Its strength is turning a hierarchy argument into comparative evidence, and it exposes the specific trap where a big element wins attention but loses comprehension — the loud number people misread as the important one. It is the test that keeps size-driven emphasis accountable to what readers actually understand.
Its failure mode is that preference alone rewards novelty or spectacle, and a single polished option invites confirmation bias — everyone nods at the one mockup they were shown. The classic misuse is presenting one hero layout and treating the approving nods as validation. The guarding discipline is to pair preference with task and accessibility outcomes and to always compare matched alternatives rather than defend a favourite.
How it implements the components¶
hierarchy_and_emphasis_map— it validates which size distribution produces the intended importance ranking and reading order, testing size contrast against real comprehension.comparative_mockup_and_multiscale_test— it builds matched paired alternatives and gathers comparative evidence; this is the comparative-mockup facet of the component.
It does not test whether real bodies physically fit or function (human_body_and_ergonomic_anchor — that is Ergonomic Fit and Clearance Trial), nor whether an engineered apparent scale reads as intended (exception_and_exaggeration_register — that is Forced-Perspective and Emphasis Test).
Related¶
- Instantiates: Proportion / Scale Calibration — this test supplies the comparative evidence for which size distribution best carries the intended hierarchy.
- Consumes: Ratio Ladder and Modular Scale supplies the size options from which the competing distributions are drawn.
- Sibling mechanisms: Cross-Medium Scale Normalization · Ergonomic Fit and Clearance Trial · Forced-Perspective and Emphasis Test · Multiscale Prototype Review · Parametric Dimension-Constraint Model · Ratio Ladder and Modular Scale · Reference-Object and Body-Scale Overlay · Responsive Typographic and Interface Scale · Scale-Drift and Exception Audit
Editorial Notes¶
Form Classification¶
Form family: Experiment, Test & Rehearsal
Rationale: Visual-Weight Mockup Comparison operates as an active test, trial, simulation, drill, or rehearsal that generates evidence through a deliberate attempt or perturbation because it compares alternative size distributions while controlling color, position, weight, and content as far as practical.
Independent corroboration: The frozen evidence defines Visual-Weight Mockup Comparison as 'Compares alternative size distributions while controlling color, position, weight, and content as far as practical', so its operative form is Experiment, Test & Rehearsal.
Nearest alternative: Assessment, Review & Assurance — Visual-Weight Mockup Comparison includes features of a bounded evaluation of existing evidence or work that produces a finding or disposition, but its defining operation is an active test, trial, simulation, drill, or rehearsal that generates evidence through a deliberate attempt or perturbation.
Review outcome: Independent reviewer agreement; medium confidence.
Origin Attribution¶
Primary origin: Art & Aesthetics
Origin pattern: Single lineage
Present-day reach: Universal
Rationale: Cleveland and McGill, Graphical Perception documents that empirical visual research compares how accurately people decode position, length, angle, area, and other graphical encodings. This is direct, mechanism-specific evidence for art aesthetics as the best-evidenced historical home of the operation—Compares alternative size distributions while controlling color, position, weight, and content as far as practical.—rather than evidence merely that the operation is useful there. The retained alternates record genuine adjacent lineages; later portability is represented separately by domain_reach=universal.
Related originating lineages:
- Human-Computer Interaction — Human Computer Interaction supplies a historically relevant adjacent lineage or formative practice for the operation—Compares alternative size distributions while controlling color, position, weight, and content as far as practical.—but the adjudicated evidence more directly locates the defining lineage in art aesthetics.
- Psychology — Psychology's perception, cognition, behavior, and risk-communication tradition contributes a separate formative lineage to the mechanism's visual weight mockup comparison logic.
- Statistics & Experimental Design — Statistics, experimental design, and measurement theory supplies a parallel or contributing lineage for the mechanism's defining operation: compares alternative size distributions while controlling color, position, weight, and content as far as practical.
Review resolution: The blind reviewers disagree on primary lineage (human_computer_interaction versus art_aesthetics). The defining operation is: Compares alternative size distributions while controlling color, position, weight, and content as far as practical. The researched Cleveland and McGill, Graphical Perception establishes that empirical visual research compares how accurately people decode position, length, angle, area, and other graphical encodings. That source therefore supports art aesthetics as the historical origin. human computer interaction remains in the uncapped alternates where it contributes a formative practice, but application or governance is not itself proof of origin. origin_mode=single_lineage records lineage construction; domain_reach=universal separately records later applicability.
Encyclopedia synthesis: The exact catalogued form synthesizes established practice rather than reproducing a single standard historical label.
Review outcome: Researched adjudication after independent review; high confidence.
Sources consulted:
Notes¶
One caveat the controlled setup can hide: the test isolates size by holding colour, position, and weight fixed, but in the shipped design those cues combine. A distribution that wins in isolation can be overpowered by a strong colour or a top-left position in production, so the winning size assignment is a starting point to re-check in the full composition — not a guarantee that survives every other cue.
[n1] Preattentive processing — the near-instant, pre-conscious visual detection of certain features, size prominent among them, studied in feature-integration research on early vision. It is why a larger element captures the eye before any reading happens, and therefore why size distribution must be tested for comprehension rather than assumed to convey importance correctly. ↩