Longitudinal Fit, Equity, and Burden Audit¶
Test or assessment — instantiates Human-Capacity Accommodation Design
Samples outcome, effort, abandonment, error, disclosure, stigma, support load, failure, repair, and disparities over time.
An accommodation that passed every test on launch day can rot in slow motion — the alternate path lags a system update, the coordinating burden creeps back onto the user, a disparity opens up that no single incident reveals. Longitudinal Fit, Equity, and Burden Audit watches the accommodation in real use, over time, sampling meaningful outcomes, effort, abandonment, error, forced disclosure, stigma, support load, failure, and repair — and slicing them across groups and modalities to expose inequity and decay. Its one defining idea is observed drift across a population and a timeline: it does not stage a failure or grade an option in a lab, it accumulates real-world signal and asks whether fit is holding, for whom, and at what hidden cost. When it finds decay or disparity, it connects the finding to a revision owner and a response threshold so monitoring triggers change rather than merely documenting harm.
Example¶
A city's paratransit service was certified accessible at rollout: accessible booking, guaranteed pickup windows, a complaint line. A year in, the audit samples the reality. Pulling twelve months of trip records and rider reports and slicing them, it finds what no single complaint showed: on-time performance is fine downtown but has quietly decayed in two outlying zones, where riders now pad an extra forty minutes; abandonment among riders who use the phone-booking channel is climbing because a website update broke the accessible booking form; and a disproportionate share of the coordination burden — re-confirming trips, chasing missed pickups — has shifted back onto riders and their families. None of these is a dramatic failure; each is a slow divergence visible only across time and groups. The audit's output routes each finding to an owner with a threshold — restore the accessible form now, investigate the zone disparity — so the drift becomes a funded correction rather than a year of quietly worse service.
How it works¶
The defining move is sampling real use across time and population, not a one-shot test:
- Instrument meaningful signal. Outcome, effort, abandonment, error, disclosure, stigma, support load, failure, and repair time — the lived cost, not request volume.
- Slice for disparity. Compare across groups, modalities, and geography to reveal who waits longer, discloses more, or is quietly worse served.
- Watch for decay against baseline. Track parity after each system change, since alternate paths tend to lag primary-system updates.
- Sample at minimum granularity. Collect the least intrusive data that answers the question, so monitoring does not slide into surveillance.
- Bind findings to a trigger. Attach every finding to a revision owner and a response threshold, or the audit merely records harm.
Tuning parameters¶
- Sampling cadence — continuous versus periodic review; frequent sampling catches decay early but raises data burden and intrusion risk.
- Disparity resolution — how finely outcomes are sliced across groups and modes; finer slices reveal hidden inequity but risk small-sample noise and re-identification.
- Signal breadth — narrow outcome metrics versus the full effort/stigma/burden slate; breadth catches invisible costs but adds collection load.
- Response threshold — how large a decay or gap triggers action; a sensitive threshold catches drift early but risks over-reacting to noise.
When it helps, and when it misleads¶
Its strength is revealing the slow, distributed failures that no single test or incident exposes: the creeping parity gap after a release, the burden quietly shifted back onto users, the disparity that only appears when outcomes are sliced by group.
Its failure mode is twofold. Left un-triggered, monitoring becomes a well-documented record of harm nobody acts on. Turned into a target, its metrics invite Goodhart's law[1] — once a measure becomes the goal, it stops measuring the thing, as staff optimize the reported number rather than real fit. The classic misuse is letting the audit metastasize into surveillance of the accommodated person instead of the system. The guarding discipline is to measure the system's fit at minimum granularity, bind each finding to an owner and threshold, and separate functional-need data from performance management, so the audit drives correction without policing the person.
How it implements the components¶
This mechanism fills the ongoing-monitoring component and nothing upstream of it:
in_context_fit_breakdown_and_revision_monitor— its entire function: sampling real-use outcome, effort, breakdown, disparity, and burden over time, and routing decay findings to a revision owner and response threshold.
It does not build capacity or demand profiles or the initial mismatch ledger (representative_human_capacity_and_variability_profile, task_environment_and_interface_demand_map, mismatch_barrier_workaround_and_risk_ledger), frame essentials (essential_outcome_user_and_context_frame), generate options (accommodation_option_and_equivalent_path_set), or grade a single option's launch-day equivalence (safety_dignity_privacy_and_burden_gate). Its nearest twin is the Accommodation Failure and Recovery Rehearsal: that one *stages a failure before dependence to test backup and repair (accommodation_implementation_support_and_ownership_plan), whereas this one observes real use over months to detect drift and disparity that no staged test reveals.*
Related¶
- Instantiates: Human-Capacity Accommodation Design — closes the lifecycle by turning observed drift into a funded revision.
- Consumes: Accommodation Failure and Recovery Rehearsal — the deployed, recovery-hardened accommodation whose in-use fit it then tracks.
- Sibling mechanisms: Person–Task–Environment Mismatch Analysis · Essential-Function and Method-Separation Review · Participatory Accommodation-Option Workshop · Multimodal Equivalence and Assistive-Compatibility Test · Accommodation Failure and Recovery Rehearsal
Editorial Notes¶
Form Classification¶
Form family: Monitoring, Sensing & Alerting
Rationale: Longitudinal Fit, Equity, and Burden Audit operates as an ongoing sensing arrangement that repeatedly observes actual state and surfaces changes or alerts because it samples outcome, effort, abandonment, error, disclosure, stigma, support load, failure, repair, and disparities over time.
Independent corroboration: The frozen evidence defines Longitudinal Fit, Equity, and Burden Audit as 'Samples outcome, effort, abandonment, error, disclosure, stigma, support load, failure, repair, and disparities over time', so its operative form is Monitoring, Sensing & Alerting.
Review outcome: Independent reviewer agreement; medium confidence.
Origin Attribution¶
Primary origin: Public Administration & Policy
Origin pattern: Cross-disciplinary synthesis
Present-day reach: Multi-domain
Rationale: Repeated review of implementation fit, equity, and burden is a policy evaluation and accountability practice.
Related originating lineages:
- Ethnography & Qualitative Methods — Lived-experience inquiry materially shapes measurement of disclosure, stigma, abandonment, and repair burden.
- Human-Computer Interaction — The audit's focus on long-run human fit, effort, error, and abandonment is rooted in human-centered design and HCI evaluation.
- Sociology & Anthropology — Distributional burden and lived institutional effects materially shape the equity inquiry.
- Statistics & Experimental Design — Longitudinal measurement and subgroup trend comparison materially support credible findings.
- Ethics of Technology & AI Governance — Equity and burden governance shape the non-performance harms included in the audit.
Review resolution: Light authoritative research supports public_administration_policy as the primary provenance: Repeated review of implementation fit, equity, and burden is a policy evaluation and accountability practice. GAO treats repeated program evaluation as an accountability and policy-management practice used to assess effectiveness and value. The competing reviewed lineage (human_computer_interaction) and other formative traditions remain explicit alternates rather than being erased or confused with downstream applicability. origin_mode=cross_disciplinary_synthesis records the relationship among those origin traditions, while domain_reach=multi_domain separately records how broadly the generalized mechanism can be applied.
Attribution caveat: The bundled longitudinal equity-and-burden instrument is an encyclopedia synthesis across HCI, qualitative research, statistics, and ethics. The bundled audit appears to be an encyclopedia synthesis across policy evaluation and equity research.
Encyclopedia synthesis: The exact catalogued form synthesizes established practice rather than reproducing a single standard historical label.
Review outcome: Researched adjudication after independent review; medium confidence.
Sources consulted:
- https://www.gao.gov/products/gao-13-570 — GAO treats repeated program evaluation as an accountability and policy-management practice used to assess effectiveness and value.
References¶
[1] Goodhart, C. A. E. "Problems of Monetary Management: The U.K. Experience". Papers in Monetary Economics, Vol. I. Reserve Bank of Australia (1976). States the Goodhart principle that a statistical regularity can collapse when used for control. registry ↩