Glossary

Searchable terminology from accessibility, web standards, and related fields.

5 results found in Speech Recognition.

Acoustic Model (AM)
An acoustic model is the component of an automatic speech recognition (ASR) system that maps short segments of audio (typically 10–25 ms frames of spectral features) to the linguistic units that produ…
Connected Speech Recognition (Continuous Speech Recognition)
A form of automatic speech recognition in which users speak words naturally, with normal coarticulation and minimal pauses, rather than pausing between each word as required by older 'discrete' or 'is…
Deaf-Accented Speech (Deaf Accent, Deaf-Accented English)
Speech produced by Deaf or Hard of Hearing people whose articulation, prosody, and voicing patterns differ from typical hearing speakers because the speaker has limited or no auditory feedback for the…
Jitter and Shimmer (Voice perturbation measures, Cycle-to-cycle variability)
Acoustic measures of voice quality that capture short-term irregularity in the vocal fold vibration. Jitter is the cycle-to-cycle variability in pitch (fundamental frequency), while shimmer is the cyc…
Speaker Adaptation (Voice Adaptation, Speaker-Adaptive Training, Voice Personalization)
Speaker adaptation is the process of adjusting an existing automatic speech recognition (ASR) system — usually one trained on a large, demographically broad corpus of able-bodied speakers — to a parti…