Literature Reviews
Reviewed research papers, articles, and publications relevant to digital accessibility.
46 results found tagged automatic speech recognition.
-
Usability Evaluation of Captions for People Who Are Deaf or Hard of Hearing
This is a SIGACCESS Newsletter article summarizing a line of research by Kafle and Huenerfauth on building a caption-quality evaluation metric that actually reflects the experience of Deaf and Hard-of-Hearing (DHH) readers — rather than simply counti…
-
Remotely Co-Designing Features for Communication Applications using Automatic Captioning with Deaf and Hearing Pairs
This CHI 2022 paper addresses two intertwined problems. First, methodologically, how can co-design research involving both Deaf/Hard-of-Hearing (DHH) and hearing participants be conducted remotely during and beyond COVID-19, when in-person sessions a…
-
Deaf Individuals' Views on Speaking Behaviors of Hearing Peers when Using an Automatic Captioning App
This CHI 2020 Late-Breaking Work paper investigates what behaviors hearing speakers should ideally exhibit when holding in-person conversations with Deaf or deaf people using an Automatic Speech Recognition (ASR) captioning app on a mobile device. Th…
-
Comparing speaker-dependent and speaker-adaptive acoustic models for recognizing dysarthric speech
This short ASSETS 2007 poster from Frank Rudzicz at the University of Toronto compares two strategies for building automatic speech recognition (ASR) acoustic models that work for people with dysarthria — a set of motor speech disorders that produces…
-
Online Quality Control for Real-Time Crowd Captioning
This paper addresses quality control in Legion:Scribe, a system that provides real-time captioning by having multiple non-expert crowd workers simultaneously type what they hear, then automatically merging their partial transcriptions into a single c…
-
Leveraging Complementary Contributions of Different Workers for Efficient Crowdsourcing of Video Captions
This paper presents BandCaption, a crowdsourcing system that combines automatic speech recognition (ASR) with input from diverse crowd workers to efficiently correct video captions. The key insight is that different groups of people — hearing-impaire…
-
From User Perceptions to Technical Improvement: Enabling People Who Stutter to Better Use Speech Recognition
This paper investigates how people who stutter (PWS) experience consumer speech recognition systems and demonstrates technical improvements that can significantly reduce errors. The work combines user research with engineering interventions across th…
-
Automatic Assessment of Speech Capability Loss in Disordered Speech
This paper investigates whether the Goodness of Pronunciation (GOP) algorithm, originally developed for computer-assisted language learning to detect non-native speaker mispronunciations, can be repurposed to assess speech capability loss in people w…
-
Perspectives on Speech and Language Interaction for Daily Assistive Technology: Introduction to Part 1 of the Special Issue
This editorial introduces the first part of a TACCESS special issue on speech and language interaction for daily assistive technology, emerging from the 2013 SLPAT (Speech and Language Processing for Assistive Technologies) workshop. The editors fram…
-
A Longitudinal Evaluation of Tablet-Based Child Speech Therapy with Apraxia World
This paper presents Apraxia World, a tablet-based speech therapy game designed for long-term home practice by children with speech sound disorders (SSDs), particularly childhood apraxia of speech (CAS). Unlike many therapy games that use simple arcad…