skip to content

A year after launch, how do you answer 'is this accessible?' with evidence rather than opinion?

level: principalimportance: should knowfreq 37%

answer

  1. Refuse the yes, answer in three parts
  2. A reading today is not a record
  3. Ask how each regression re-entered
  4. Name what nobody has ever walked
  5. Completable is not the same as conformant

basics

~20 s

Answer with evidence: dated records of which task was tested, by which pass, by whom, and what was found; a gate blocking regressions on every change; known unfixed defects with owners; and findings from people with disabilities.

solid answer

~50 s

Refuse the yes-or-no and answer in three parts. **What was tested** — which tasks, by which passes, on which dates, by whom, and what was found; a scan run this morning is a reading, not a record. **What stops it decaying** — the measurable criteria under a gate on every change, per-control guarantees at the shared library boundary, and manual passes attached to changes rather than to a quarterly phase; every regression that re-entered should have a recorded cause. **What you still do not know** — the tasks nobody has walked, the screens nobody will fix, and the criteria you have never evaluated, each with an owner and a date. Then say what involving people with disabilities changed: it moves the question from whether criteria were met to whether the task was completable, and those two answers routinely differ.

go deeper

for a junior

Be ready to say that the honest answer depends on what was tested and when, and that a tool result from today does not by itself show a product is accessible.

for a middle

Explain what a durable record contains — the task, the pass, the date, the tester, the findings — and why that beats a scan result when somebody asks the question months later.

for a senior

Show how you stop decay: which criteria are gated on every change, where per-control guarantees live, and the habit of asking, for every regression, which mechanism should have caught it.

for a principal

Own the whole answer: the budget split across gating, auditing and participation, the risk you accept explicitly on screens nobody will fix, and why counts of findings are a corrupting target.

## The question is a trap, and the answer is a portfolio "Is this accessible?" invites a yes, and a yes a year after launch is almost always built on the wrong thing — a scan run this morning, or a memory of an effort that finished eleven months ago. The credible answer has three parts, and a lead's job is to make sure the organisation can produce all three on demand. 1. **Evidence of what was tested.** 2. **Evidence that it has not decayed.** 3. **An honest statement of what is still unknown.** ## Part one: what was actually tested A reading is what a tool says today. A **record** is what somebody did, when, and what they found. The record that holds up a year later has, per entry: the **task** (not the screen), the **pass** used, the **date**, **who** ran it, the **findings**, and the **decision** on each finding. Two properties make it credible: - It is scoped by task, so it maps onto what a user was trying to do rather than onto a route. - It includes the failures. A record with no findings anywhere is evidence that the passes were not really run. Alongside it sits the **defect log**, with each entry labelled by the requirement it breaks in words — "the announced name does not match the visible label" — so the log stays readable when personnel change. ## Part two: what stops it decaying Accessibility decays by default, because every change is an opportunity to remove a name or reorder a screen. Four mechanisms, and the mix is a real budget decision: | Mechanism | What it protects | Cost profile | |---|---|---| | Automated gate on every change | The measurable criteria, continuously | High setup, near-zero marginal | | Guarantees at the shared library boundary | Per-control naming, role and operability, across every consumer | Moderate, and it propagates | | Manual passes attached to changes | Meaning-level criteria on the surface that moved | Linear in change volume | | Periodic deeper pass over whole tasks | Composition and real paths | Lumpy, and blind between runs | The measurement that matters here is not the count of open findings. It is **how defects re-enter**. For every regression, ask which mechanism should have caught it and did not: a naming regression that a gate could have failed means the gate is mis-scoped; a broken order that only the quarterly pass found means the manual pass is not attached to changes. A year of that question, answered honestly, is worth more than a year of violation counts, which fall whenever rules are disabled. ## Part three: what you still do not know The part that separates a credible answer from a sales answer. Name, with an owner and a date each: - **Tasks nobody has walked.** Coverage by task, stated as a fraction, is far more honest than coverage by screen. - **Criteria never evaluated.** Most of the standard is not in anybody's rule set, so if no human ever judged them, say so. - **Known and unfixed defects.** A maintenance console for farm machinery has a 6-step application form whose step 4 is a legacy screen nobody will touch. After 14 months there are 4 open findings on it, and a lead's obligation is to make that an explicit, owned decision — with the risk stated and a route around the screen offered — rather than a silence that reads as a pass. ## What involving people with disabilities changes It changes the question being answered. Criteria-based testing asks whether requirements were met; a session with somebody who uses an assistive technology daily asks whether the **task was completed**, and the two answers diverge constantly. Three things it reliably surfaces that no pass does: - **Cost, not just possibility.** A task that takes 40 steps and 3 wrong turns satisfies the criteria and is abandoned in practice. - **Priority.** Practitioners rank findings differently from teams; the defect an engineer calls cosmetic is often the one that ends the session. - **Paths nobody designed.** Real users arrive mid-flow, with settings the team never tried. Two cautions a lead owns. First, participation is **paid, consented work with real practitioners**, not a favour extracted from a colleague. Second, it does not replace criteria-based testing: a handful of sessions cannot cover the standard, and criteria coverage cannot tell you a task is usable. Each answers what the other cannot. ## Saying it out loud "Of the 12 core tasks, 9 have a dated keyboard and assistive-technology pass in the last two quarters. The measurable criteria are gated on every change and have blocked 7 regressions this year. Two practitioner sessions changed our ordering of the backlog. We have 4 known open defects on one legacy screen, owned and risk-accepted, and 3 tasks nobody has walked yet." That is an answer with a shape somebody can audit — and it is the shape to give even when parts of it are uncomfortable.

  • Your open-findings count has halved this year. Why might that not be good news?
    A count falls for two very different reasons: defects were fixed, or findings stopped being produced. Rules get disabled, needs-review items get bulk-closed, manual passes quietly stop happening. Before celebrating, check that the number of passes run held steady and that the mix of finding types did not change. Counts are a reporting aid, never a goal to manage toward.
  • How do you decide between more internal testing and bringing in an external assessor?
    Buy the thing you cannot produce yourself. More internal testing scales what the team already knows how to look for; an external assessor brings independence, breadth across the criteria, and a reading nobody on the team has an incentive to soften. Reach for the assessor before a major commitment or after a redesign, and spend the rest on continuous protection, which an assessment cannot provide.
  • One legacy screen has open defects nobody will fix. What is the lead's obligation?
    Make the decision explicit and owned rather than letting silence stand in for a pass. Record the defects, the tasks they block, the severity and the accepted risk, name who accepted it, and revisit on a date. Where possible, offer a route around the screen. The unacceptable outcome is that the next audit rediscovers them as though they were news.

saying these in an interview costs you the question

  • Answers yes on the strength of a scan run today
  • Keeps no record of which tasks were tested or by whom
  • Treats an absence of complaints as evidence of accessibility
  • Consults people with disabilities once, at the very end
  • Fixes each regression without asking how it re-entered