Is accessibility conformance an elusive property? A study of validity and reliability of WCAG 2.0-Reference-Cited by-同舟云学术

Is accessibility conformance an elusive property? A study of validity and reliability of WCAG 2.0

Published:2012-03 Issue:2 Volume:4 Page:1-28
ISSN:1936-7228
Container-title:ACM Transactions on Accessible Computing
language:en
Short-container-title:ACM Trans. Access. Comput.

Author:

Brajnik Giorgio¹,Yesilada Yeliz²,Harper Simon³

Affiliation:

1. University of Udine, Italy

2. Middle East Technical University, Northern Cyprus Campus, Turkey

3. University of Manchester, UK

Abstract

The Web Content Accessibility Guidelines (WCAG) 2.0 separate testing into both “Machine” and “Human” audits; and further classify “Human Testability” into “Reliably Human Testable” and “Not Reliably Testable”; it is human testability that is the focus of this paper. We wanted to investigate the likelihood that “at least 80% of knowledgeable human evaluators would agree on the conclusion” of an accessibility audit, and therefore understand the percentage of success criteria that could be described as reliably human testable, and those that could not. In this case, we recruited twenty-five experienced evaluators to audit four pages for WCAG 2.0 conformance. These pages were chosen to differ in layout, complexity, and accessibility support, thereby creating a small but variable sample. We found that an 80% agreement between experienced evaluators almost never occurred and that the average agreement was at the 70--75% mark, while the error rate was around 29%. Further, trained—but novice—evaluators performing the same audits exhibited the same agreement to that of our more experienced ones, but a reduction on validity of 6--13% ; the validity that an untrained user would attain can only be a conjecture. Expertise appears to improve (by 19%) the ability to avoid false positives. Finally, pooling the results of two independent experienced evaluators would be the best option, capturing at most 76% of the true problems and producing only 24% of false positives. Any other independent combination of audits would achieve worse results. This means that an 80% target for agreement, when audits are conducted without communication between evaluators, is not attainable, even with experienced evaluators, when working on pages similar to the ones used in this experiment; that the error rate even for experienced evaluators is relatively high and further, that untrained accessibility auditors be they developers or quality testers from other domains, would do much worse than this.

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Science Applications,Human-Computer Interaction

Link

https://dl.acm.org/doi/pdf/10.1145/2141943.2141946

Reference36 articles.

1. On the testability of WCAG 2.0 for beginners

2. Comparing accessibility evaluation tools: a method for tool effectiveness

3. Beyond Conformance: The Role of Accessibility Evaluation Methods

4. Testability and validity of WCAG 2.0

Cited by 27 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Turning manual web accessibility success criteria into automatic: an LLM-based approach;Universal Access in the Information Society;2024-03-16

2. Applying andragogy for integrating a MOOC into a formal online learning experience in computer engineering;Heliyon;2024-01

3. Automating Error Identification and Evaluating Web Accessibility for Differently Abled Users;Lecture Notes in Networks and Systems;2024

4. The Transparency of Automatic Web Accessibility Evaluation Tools: Design Criteria, State of the Art, and User Perception;ACM Transactions on Accessible Computing;2023-03-28

5. A Case Study to Explore a UDL Evaluation Framework Based on MOOCs;Applied Sciences;2022-12-29