Educational companion dossier · Fact, interpretation, lived experience, clinical education, fiction, and mechanics are labeled separately. Scope & safety
REAL-WORLD INTERPRETIVE

EVIDENCE REVIEW WORKBENCHES

Evaluation Evidence, Corrections, and Re-Test Records

A version-bound method for recording what was tested, what failed, what changed, who owned the remedy, and which evidence must be repeated before a route or revision can be relied upon.

LEVEL 1

ORIENTATION

Why this matters

REAL-WORLD INTERPRETIVE

One-sentence brief

A usability or accessibility finding has little value if it cannot be reproduced, tied to an exact revision, propagated to dependent content, and re-tested after correction.

REAL-WORLD INTERPRETIVE

Three key points

  1. Bind every observation to an exact route, revision, configuration, task, and access mode.
  2. Separate raw participant evidence, analysis, public summary, correction decision, and re-test result.
  3. Reopen affected checks whenever a material source, interface, language, access path, rule, or authority boundary changes.
LEVEL 2

WORKING BRIEF

Evidence, context, and limits

REAL-WORLD INTERPRETIVE

Bind evidence to the exact revision

Record the release, route, content or interface fingerprint, browser and device class, language, access mode, test fixture, and date. A result for one revision does not silently transfer to a changed one.

REAL-WORLD INTERPRETIVE

Separate observation from interpretation

Observed task events, participant statements, assistive-technology output, timing, and errors should be distinguishable from the evaluator's interpretation, severity, and proposed remedy.

REAL-WORLD INTERPRETIVE

Use stable issue and rule identifiers

A material issue needs a stable identifier, reproduction steps, affected evidence or task, severity rationale, owner, status, and links to every route or record that may inherit the defect.

REAL-WORLD INTERPRETIVE

Document the correction and its limits

The correction record states what changed, what did not change, which risks remain, whether static parity or rollback was used, and which previous evidence is superseded rather than deleted.

REAL-WORLD INTERPRETIVE

Repeat the right evidence

Re-test the failed task and the nearby regression risks, including access paths and low-bandwidth or reduced-motion alternatives. Automated regression can confirm code behavior but cannot substitute for human task evidence where the defect was human-facing.

REAL-WORLD INTERPRETIVE

Publish a minimized, non-identifying summary

Public correction notes can describe the affected route, issue, remedy, date, and remaining limitation without exposing participant identity, sensitive accommodations, raw recordings, or private review details.

LEVEL 3

COMPLETE DOSSIER

Limitations, game links, and review context

DISPUTED / MULTIPLE ACCOUNTS

Known limitations and gaps

  • The WIP.25 material plans evaluation and accountable review; it does not claim that representative users, assistive-technology users, specialists, affected communities, or regulators have completed a review.
  • The preserved design reports remain unreviewed research leads. Their named institutions, standards, thresholds, algorithms, and performance claims require claim-specific verification before external reliance.
  • The local-only planner accepts bounded declarations rather than live participant data, recordings, free text, uploads, credentials, or personal profiles.
  • No result is accessibility certification, legal advice, specialist approval, publication authority, activation authority, production readiness, or permission to tune a live system.
  • Protected traits, diagnosis, disability, language, nationality, religion, migration history, poverty, and other vulnerable status are never defects, risk scores, quality rankings, or exclusion criteria.
REAL-WORLD INTERPRETIVE

Decision matrix

Evaluation evidence lifecycle.
Stage Required record Authority limit Reopening trigger
Plan Task, coverage, comparator, measures, privacy, stop rule Not evidence that testing occurred Material scope or risk change
Observe Version-bound task and access-mode evidence Observation is not final interpretation Reproduction failure or missing context
Decide Scoped severity, owner, remedy, rollback, exclusions Decision applies only within stated authority New evidence or conflict
Correct Exact change and propagation record Change is not proof of resolution Regression or dependent-route change
Re-test Repeated failed task and adjacent regression evidence Pass is not universal certification Later material revision
REAL-WORLD INTERPRETIVE

Publication audit checklist

Revision binding

Can the evidence be tied to the exact tested route and revision?

Pass condition: Release, route, configuration, language, task, and access mode are recorded.

Reproducibility

Can another accountable evaluator reproduce the finding?

Pass condition: Observation, steps, expected behavior, actual behavior, and evidence references are distinct.

Propagation

Did the remedy reach every affected public and protected derivative?

Pass condition: Content, templates, rules, UAI records, docs, tests, sitemap, and manifests are updated together where applicable.

Re-test

Was the failed task and nearby regression risk repeated after correction?

Pass condition: The result is bound to the corrected revision and residual limitations remain visible.

LEVEL 4

RESEARCH EDITION

Sources, methods, and stable links

REAL-WORLD VERIFIED

Method and corrections

This page follows the public method for provenance, confidence, source independence, alternative accounts, limitations, review state, and visible correction.

NEXT

CONTINUE

Related learning

Evidence Review WorkbenchesAttribution Evidence Review WorkbenchA deterministic, browser-local workbench that turns a bounded public claim state into AT-* review prompts covering rung mismatch, source independence, claimant labeling, evidence custody, reach/effect separation, correction, and identity risk.Evidence Review WorkbenchesCorrection Propagation PlannerA deterministic browser-local workbench that turns bounded declarations about claim state, dependent routes, public visibility, protected memory, response, re-test, and reopening into stable CP-* review prompts.Evidence Review WorkbenchesEvidence Review Workbenches: Local Planning, Source Chains, Specialist Packets, Human Testing, and Accountable ReviewEight transparent, local-only workbenches connect lifecycle, visualization, cognitive load, representative testing, attribution, source-chain independence, correction propagation, and specialist-packet readiness without uploads, profiling, hidden scoring, or implied authority.Evidence Review WorkbenchesRepresentative Testing Protocol PlannerA deterministic, browser-local planning aid for defining a bounded task, representative participant coverage, assistive-technology checks, comparison conditions, outcome evidence, correction, and re-test triggers. Learning pathAccountable User Testing and Specialist ReviewMove from a transparent heuristic through representative task design, assistive-technology coverage, evidence capture, correction, re-test, and scoped specialist disposition without self-approval.