Educational companion dossier · Fact, interpretation, lived experience, clinical education, fiction, and mechanics are labeled separately. Scope & safety
REAL-WORLD INTERPRETIVE

EVIDENCE REVIEW WORKBENCHES

Representative Testing Protocol Planner

A deterministic, browser-local planning aid for defining a bounded task, representative participant coverage, assistive-technology checks, comparison conditions, outcome evidence, correction, and re-test triggers.

LEVEL 1

ORIENTATION

Why this matters

REAL-WORLD INTERPRETIVE

One-sentence brief

An interface can pass automated checks and still fail real tasks, assistive technologies, language contexts, devices, or recovery scenarios. A protocol makes the missing evidence explicit before anyone mistakes a prototype for a validated experience.

REAL-WORLD INTERPRETIVE

Three key points

  1. Define the task, comparison condition, participant coverage, access modes, device context, outcomes, and re-test trigger before running a study.
  2. Measure task success, comprehension, error recovery, and source inspection rather than engagement or satisfaction alone.
  3. Treat a protocol-ready result as a plan for accountable human execution, never as evidence that testing or approval already occurred.
LAB

LOCAL-ONLY PROTOCOL PLANNER

Plan accountable representative-user and assistive-technology testing

Protocol readiness is not testing completion

This planner uses bounded declarations to expose missing task, coverage, comparison, outcome, evidence, correction, and re-test conditions. It collects no participant data and cannot certify accessibility, record specialist approval, authorize publication, activate a system, or unblock production tuning.

Task and coverage
Evidence and accountability
LEVEL 2

WORKING BRIEF

Evidence, context, and limits

REAL-WORLD INTERPRETIVE

Start with a bounded task

A valid protocol states what a participant must locate, compare, explain, decide, or recover. Vague goals such as make it engaging cannot establish whether the interface preserved evidence, context, or access.

REAL-WORLD INTERPRETIVE

Plan representative coverage without ranking people

Recruitment should reflect the actual audience, language, expertise, device, bandwidth, disability, and assistive-technology contexts relevant to the task. Coverage identifies who must be included; it does not score a population or define a universal user.

REAL-WORLD INTERPRETIVE

Use a meaningful comparison condition

A prototype should be compared with a simpler baseline, prior version, static equivalent, or alternative presentation. Without a comparator, a fast task may still conceal preventable errors or unequal access.

REAL-WORLD INTERPRETIVE

Measure outcomes and recovery

Task success, comprehension, source tracing, error rate, correction, context recovery, and ability to stop or reverse an action provide stronger evidence than dwell time, clicks, virality, or subjective enthusiasm alone.

REAL-WORLD INTERPRETIVE

Minimize participant data

Collect only data necessary for the declared evaluation purpose. Separate consent, contact information, recordings, accessibility accommodations, raw observations, and published summaries. Do not reuse the material for profiling, secondary training, targeting, or unrelated analytics.

REAL-WORLD INTERPRETIVE

Specify correction and re-test triggers

Every material defect needs an owner, reproducible condition, affected route or revision, remedy, and re-test rule. A changed interface, source, language, device, access path, or authority boundary can reopen the protocol.

LEVEL 3

COMPLETE DOSSIER

Limitations, game links, and review context

DISPUTED / MULTIPLE ACCOUNTS

Known limitations and gaps

  • The WIP.25 material plans evaluation and accountable review; it does not claim that representative users, assistive-technology users, specialists, affected communities, or regulators have completed a review.
  • The preserved design reports remain unreviewed research leads. Their named institutions, standards, thresholds, algorithms, and performance claims require claim-specific verification before external reliance.
  • The local-only planner accepts bounded declarations rather than live participant data, recordings, free text, uploads, credentials, or personal profiles.
  • No result is accessibility certification, legal advice, specialist approval, publication authority, activation authority, production readiness, or permission to tune a live system.
  • Protected traits, diagnosis, disability, language, nationality, religion, migration history, poverty, and other vulnerable status are never defects, risk scores, quality rankings, or exclusion criteria.
REAL-WORLD INTERPRETIVE

Decision matrix

Evaluation protocol readiness matrix.
Protocol element Insufficient Review-ready declaration Evidence produced later
Task General aspiration or engagement goal Bounded action and success condition Observed completion, errors, comprehension, recovery
Participant coverage Convenience sample only Audience, language, disability, expertise, and access contexts named Coverage log and exclusions
Comparison Prototype viewed in isolation Baseline, prior version, or equivalent alternative Difference in outcomes and failure modes
Access modes Keyboard or screen reader assumed Named AT, zoom, contrast, motion, mobile, and low-bandwidth paths Task evidence for each relevant mode
Correction Issues noted informally Stable issue, owner, remedy, re-test trigger, and rollback state Closed or reopened evidence record
REAL-WORLD INTERPRETIVE

Publication audit checklist

Bounded task

Is the participant asked to complete a specific evidence or interaction task?

Pass condition: The protocol states the action, success condition, source context, and time or safety constraints.

Representative coverage

Does the protocol include the language, expertise, device, bandwidth, disability, and assistive-technology contexts that can materially change the task?

Pass condition: Included and excluded contexts are explicit; no group is treated as a quality score.

Outcome evidence

Will the study capture success, comprehension, errors, correction, source inspection, and recovery?

Pass condition: Engagement and satisfaction are supplementary rather than substitutes for task evidence.

Correction and re-test

Can a material finding be reproduced, assigned, corrected, and re-tested against the same revision?

Pass condition: Stable identifiers, version binding, remedy, rollback, and reopening conditions are documented.

LEVEL 4

RESEARCH EDITION

Sources, methods, and stable links

REAL-WORLD VERIFIED

Method and corrections

This page follows the public method for provenance, confidence, source independence, alternative accounts, limitations, review state, and visible correction.

NEXT

CONTINUE

Related learning

Evidence Review WorkbenchesAttribution Evidence Review WorkbenchA deterministic, browser-local workbench that turns a bounded public claim state into AT-* review prompts covering rung mismatch, source independence, claimant labeling, evidence custody, reach/effect separation, correction, and identity risk.Evidence Review WorkbenchesCorrection Propagation PlannerA deterministic browser-local workbench that turns bounded declarations about claim state, dependent routes, public visibility, protected memory, response, re-test, and reopening into stable CP-* review prompts.Evidence Review WorkbenchesEvaluation Evidence, Corrections, and Re-Test RecordsA version-bound method for recording what was tested, what failed, what changed, who owned the remedy, and which evidence must be repeated before a route or revision can be relied upon.Evidence Review WorkbenchesEvidence Review Workbenches: Local Planning, Source Chains, Specialist Packets, Human Testing, and Accountable ReviewEight transparent, local-only workbenches connect lifecycle, visualization, cognitive load, representative testing, attribution, source-chain independence, correction propagation, and specialist-packet readiness without uploads, profiling, hidden scoring, or implied authority. Learning pathAccountable User Testing and Specialist ReviewMove from a transparent heuristic through representative task design, assistive-technology coverage, evidence capture, correction, re-test, and scoped specialist disposition without self-approval.