Literature Reviews
Reviewed research papers, articles, and publications relevant to digital accessibility.
77 results found tagged speech recognition.
-
Capti-Speak: A Speech-Enabled Web Screen Reader
This paper presents Capti-Speak, a speech-augmented screen reader for web browsing that allows blind users to combine natural language voice commands with traditional keyboard shortcuts. Built as an extension to the Capti Narrator screen reader, Capt…
-
Automatically Generating and Improving Voice Command Interface from Operation Sequences on Smartphones
This paper presents AutoVCI, a system that automatically generates voice command interfaces (VCIs) for smartphone tasks from recorded touch operation sequences, enabling hands-free and eyes-free interaction without requiring programming expertise, co…
-
Evaluation of Real-time Captioning by Machine Recognition with Human Support
This paper from IBM Research Tokyo investigates a hybrid approach to real-time captioning that combines Automated Speech Recognition (ASR) with human correction to make workplace meetings accessible for deaf and hard of hearing (DHH) employees. Profe…
-
The Effects of Automatic Speech Recognition Quality on Human Transcription Latency
This paper from Carnegie Mellon University and the University of Michigan empirically investigates when automatic speech recognition (ASR) output helps or hinders human transcriptionists producing captions for deaf and hard of hearing people. Manual …
-
WebReader: a screen reader for everyone, everywhere
This extended abstract presents WebReader, a free and open source JavaScript library that implements a subset of screen reader features directly within web pages, requiring no software installation beyond a web browser. The project addresses two key …
-
Vocal Programming for People with Upper-Body Motor Impairments
This paper presents VocalIDE, a prototype voice-based integrated development environment (IDE) designed to enable people with upper-body motor impairments to write and edit computer code using speech commands rather than a keyboard. Only 4% of profes…
-
Exploration of Automatic Speech Recognition for Deaf and Hard of Hearing Students in Higher Education Classes
This paper presents a qualitative study of how deaf and hard of hearing (DHH) students at the National Technical Institute for the Deaf (Rochester Institute of Technology) experienced automatic speech recognition (ASR) as a supplemental access servic…
-
Automatic Generation and Evaluation of Usable and Secure Audio reCAPTCHA
This paper presents reCAPGen, a system that automatically generates usable and secure audio CAPTCHAs by leveraging the gap between human and machine speech recognition abilities. Visual CAPTCHAs — the dominant form of online human verification — are …
-
Handsfree for Web: A Google Chrome extension to browse the web via voice commands
This demonstration paper presents Handsfree for Web, a free Google Chrome browser extension that enables users to browse the web entirely through voice commands. The tool addresses a fundamental accessibility barrier: nearly all websites require manu…
-
Pushpak: Voice Command-based eBook Navigator
This demonstration paper presents Pushpak, a voice command-based eBook navigator designed to reduce the steep learning curve associated with screen reader software. The authors identify a key accessibility barrier: effective use of screen readers lik…