Glossary
Searchable terminology from accessibility, web standards, and related fields.
55 results found in Evaluation Methods.
- AI Auditing (Algorithmic Auditing, AI Audit)
- The systematic evaluation of an AI system's outputs, behaviour, or training data to identify harms such as bias, stereotype reproduction, or accessibility failures. Audits may be conducted by industry…
- Accessibility Evaluation Method (AEM, Accessibility Testing Method)
- A structured approach or procedure used to assess the accessibility of digital products, websites, or applications. Accessibility evaluation methods include conformance review (checking against standa…
- Accessibility Heuristics (Accessibility Heuristic Evaluation)
- A set of broad usability and accessibility principles used to evaluate digital products for barriers that may prevent people with disabilities from using them effectively. Unlike detailed technical ch…
- Accessibility Inspection (Accessibility Inspection Method, Accessibility Audit)
- An evaluation approach in which an expert or designer reviews an interface against a set of accessibility criteria without recruiting end users, analogous to usability inspection methods such as heuri…
- BLEU Score (BiLingual Evaluation Understudy, BLEU)
- A metric for evaluating the quality of machine-generated text by comparing it to one or more reference (human-written) translations. BLEU calculates precision by counting how many n-grams (sequences o…
- Back-Translation (Reverse Translation)
- A quality assurance method used in survey and instrument translation where a translated version is independently translated back into the original language by a different translator. The back-translat…
- Barrier Walkthrough (BW Method)
- The Barrier Walkthrough is a structured expert evaluation method for assessing web accessibility in which evaluators systematically examine a website against a predefined set of accessibility barriers…
- Between-Subjects Design (Between-Groups Design, Independent-Groups Design)
- A between-subjects design is an experimental research design in which each participant is assigned to only one condition, and the conditions are compared across different groups of people. It contrast…
- Blurred Vision Simulation (Vision Simulation, Low Vision Simulation)
- A technique used in accessibility evaluation where evaluators simulate the visual experience of people with reduced visual acuity by artificially blurring their view of a website or application. Metho…
- Correctness (Precision, Validity)
- In the context of accessibility evaluation, correctness (also called precision) is the proportion of reported accessibility problems that are true problems — that is, issues that genuinely affect user…
- Counterfactual Explanation (Counterfactual XAI)
- An explanation technique that communicates what minimal change to the input would have produced a different output from an AI model, for example 'if the applicant's income had been $5,000 higher, the …
- Criterion Validity
- A psychometric property indicating whether an instrument's scores relate to some external measurable criterion. In practice, this is assessed by comparing the instrument's results with scores from ano…
- Critical Incident Questionnaire (CIQ)
- A short, open-ended reflective tool developed by Stephen Brookfield for teaching and learning contexts, typically consisting of five questions asking participants to recall moments from a recent exper…
- Cursor Deviation (Cursor Drift, Path Deviation)
- The difference between the actual path taken by a cursor and the ideal straight-line path between the starting point and the target. Cursor deviation is a key performance metric in evaluating alternat…
- Disability-Centered Evaluation (Disability-Centric Evaluation, Disability-First Evaluation)
- An approach to evaluating AI systems, tools, or research artefacts that places disabled people's lived experiences, information needs, and failure contexts at the centre of study design — including wh…
- Discriminative Ability (Discriminative ability of a metric, Discriminability)
- In accessibility research methodology, the property of an evaluation metric to reveal statistically significant differences between stimuli that are known to differ along the dimension being measured.…
- End-User Auditing (User-Led Auditing, End User Audits)
- An approach to AI auditing in which everyday users — rather than professional evaluators — identify problems, biases, or harms in AI outputs based on their lived experience. End-user auditing is parti…
- Equal Error Rate (EER, Crossover Error Rate)
- A metric used to evaluate biometric system performance, representing the point at which the false acceptance rate (wrongly accepting unauthorized users) equals the false rejection rate (wrongly reject…
- Evaluation Reliability (Inter-rater Reliability, Evaluator Agreement)
- The extent to which independent accessibility evaluations of the same content produce consistent results. High reliability means that different evaluators using the same method will identify similar s…
- F-measure (F-score, F1 Score)
- A metric that combines correctness (precision) and sensitivity (recall) into a single balanced score, calculated as the harmonic mean of the two values. In accessibility evaluation research, the F-mea…