Glossary

Searchable terminology from accessibility, web standards, and related fields.

55 results found in Evaluation Methods.

AI Auditing (Algorithmic Auditing, AI Audit)
The systematic evaluation of an AI system's outputs, behaviour, or training data to identify harms such as bias, stereotype reproduction, or accessibility failures. Audits may be conducted by industry…
Accessibility Evaluation Method (AEM, Accessibility Testing Method)
A structured approach or procedure used to assess the accessibility of digital products, websites, or applications. Accessibility evaluation methods include conformance review (checking against standa…
Accessibility Heuristics (Accessibility Heuristic Evaluation)
A set of broad usability and accessibility principles used to evaluate digital products for barriers that may prevent people with disabilities from using them effectively. Unlike detailed technical ch…
Accessibility Inspection (Accessibility Inspection Method, Accessibility Audit)
An evaluation approach in which an expert or designer reviews an interface against a set of accessibility criteria without recruiting end users, analogous to usability inspection methods such as heuri…
BLEU Score (BiLingual Evaluation Understudy, BLEU)
A metric for evaluating the quality of machine-generated text by comparing it to one or more reference (human-written) translations. BLEU calculates precision by counting how many n-grams (sequences o…
Back-Translation (Reverse Translation)
A quality assurance method used in survey and instrument translation where a translated version is independently translated back into the original language by a different translator. The back-translat…
Barrier Walkthrough (BW Method)
The Barrier Walkthrough is a structured expert evaluation method for assessing web accessibility in which evaluators systematically examine a website against a predefined set of accessibility barriers…
Between-Subjects Design (Between-Groups Design, Independent-Groups Design)
A between-subjects design is an experimental research design in which each participant is assigned to only one condition, and the conditions are compared across different groups of people. It contrast…
Blurred Vision Simulation (Vision Simulation, Low Vision Simulation)
A technique used in accessibility evaluation where evaluators simulate the visual experience of people with reduced visual acuity by artificially blurring their view of a website or application. Metho…
Correctness (Precision, Validity)
In the context of accessibility evaluation, correctness (also called precision) is the proportion of reported accessibility problems that are true problems — that is, issues that genuinely affect user…
Counterfactual Explanation (Counterfactual XAI)
An explanation technique that communicates what minimal change to the input would have produced a different output from an AI model, for example 'if the applicant's income had been $5,000 higher, the …
Criterion Validity
A psychometric property indicating whether an instrument's scores relate to some external measurable criterion. In practice, this is assessed by comparing the instrument's results with scores from ano…
Critical Incident Questionnaire (CIQ)
A short, open-ended reflective tool developed by Stephen Brookfield for teaching and learning contexts, typically consisting of five questions asking participants to recall moments from a recent exper…
Cursor Deviation (Cursor Drift, Path Deviation)
The difference between the actual path taken by a cursor and the ideal straight-line path between the starting point and the target. Cursor deviation is a key performance metric in evaluating alternat…
Disability-Centered Evaluation (Disability-Centric Evaluation, Disability-First Evaluation)
An approach to evaluating AI systems, tools, or research artefacts that places disabled people's lived experiences, information needs, and failure contexts at the centre of study design — including wh…
Discriminative Ability (Discriminative ability of a metric, Discriminability)
In accessibility research methodology, the property of an evaluation metric to reveal statistically significant differences between stimuli that are known to differ along the dimension being measured.…
End-User Auditing (User-Led Auditing, End User Audits)
An approach to AI auditing in which everyday users — rather than professional evaluators — identify problems, biases, or harms in AI outputs based on their lived experience. End-user auditing is parti…
Equal Error Rate (EER, Crossover Error Rate)
A metric used to evaluate biometric system performance, representing the point at which the false acceptance rate (wrongly accepting unauthorized users) equals the false rejection rate (wrongly reject…
Evaluation Reliability (Inter-rater Reliability, Evaluator Agreement)
The extent to which independent accessibility evaluations of the same content produce consistent results. High reliability means that different evaluators using the same method will identify similar s…
F-measure (F-score, F1 Score)
A metric that combines correctness (precision) and sensitivity (recall) into a single balanced score, calculated as the harmonic mean of the two values. In accessibility evaluation research, the F-mea…