Glossary
Searchable terminology from accessibility, web standards, and related fields.
47 results found in captioning.
- Live Captioning (Real-Time Captioning, Live Captions)
- The process of converting spoken language into text displayed in real time, enabling Deaf and hard of hearing individuals to follow live audio content such as meetings, lectures, broadcasts, and event…
- Logocentrism
- In captioning studies, the systematic prioritization of speech and spoken language over non-speech sounds in captioning practices and technologies. Logocentrism in captioning manifests as speech capti…
- Non-Speech Captions (Non-Speech Sound Captions, Non-Dialogue Captions)
- Textual descriptions of non-speech audio elements in media content, including environmental sounds, music, and sound effects, displayed as part of closed or open captions. Non-speech captions are esse…
- Non-Speech Information (NSI, Non-Dialogue Audio Information)
- Any audio content in media that is not spoken dialogue, including environmental sounds, music, sound effects, and ambient noise. Non-speech information plays a critical role in storytelling by conveyi…
- Non-Speech Sounds (Non-Speech Audio, Sound Effects)
- Auditory content in media that is not spoken dialogue, including music, environmental noises, sound effects, laughter, applause, and other ambient sounds. Non-speech sounds carry important narrative, …
- Onomatopoeia
- Words that phonetically imitate or suggest the sound they describe, such as "buzz," "crash," "swoosh," or "sizzle." In captioning, onomatopoeia is one approach to representing non-speech sounds, offer…
- Open Captioning (Open Captions, Burned-In Captions)
- Captions that are permanently embedded into the video image and cannot be turned off by the viewer. Unlike closed captions, open captions are part of the visual content itself, making them visible to …
- Open Captions (Burned-in Captions, Hard-coded Captions)
- Captions that are permanently embedded into a video and cannot be turned off by the viewer. Unlike closed captions, which can be toggled on or off, open captions are always visible as part of the vide…
- Paralinguistic Cues (Paralanguage, Paralinguistic Features, Non-verbal Vocal Cues)
- Aspects of spoken communication that carry meaning beyond the literal words themselves: tone of voice, pitch contour, loudness, rhythm, tempo, stress, pauses, and voice quality. Paralinguistic cues co…
- Play-by-Play (Play-by-play announcing, Play-by-play commentary)
- In sports broadcasting, the moment-to-moment verbal description of on-screen action provided by the main commentator (e.g., who has the puck, who is passing to whom). Because play-by-play describes wh…
- Pop-on Captions (Pop-on style, Block captions)
- A captioning display style in which a complete caption appears on screen as a single block, remains visible for a readable duration, and is then replaced in one transition by the next block. Pop-on ca…
- Rapid Serial Visual Presentation (RSVP)
- A text display method in which words or short phrases are shown one at a time in a fixed location on screen in quick succession, eliminating the need for eye movements (saccades) between words. RSVP w…
- Real-Time Captioning (Live Captioning, CART, Communication Access Realtime Translation)
- The process of converting spoken language into text display in real time, typically with only a few seconds of delay. Professional real-time captioning (CART) uses stenographers with specialised short…
- Roll-up Captions (Roll-up style, Scroll-up Captions)
- A captioning display style in which text is added one word or line at a time, scrolling upward as new text arrives and pushing earlier lines off the top. Roll-up is typically used in live captioning b…
- Shadow Speaking (Shadow Captioning, Respeaking)
- A captioning technique where a trained human operator listens to live speech and repeats (or "respeaks") it clearly into a speech recognition system, which then generates real-time captions. The shado…
- Sound Event Detection (Audio Tagging, Automatic Sound Recognition)
- A machine learning technique that automatically identifies and classifies sounds within an audio stream, such as music, applause, laughter, environmental noises, and other non-speech audio events. In …
- Sound Representation (Sound Depiction)
- The methods and conventions used to convey audio information through text in captions and other written formats. Common approaches include descriptive text (explaining the sound source and quality), o…
- Speaker Identification (Speaker ID, Speaker Attribution)
- Methods used in captions and subtitles to indicate which person is currently speaking, enabling viewers to follow conversations among multiple participants. Common in-text speaker identification techn…
- Speech-modulated Typography (Speech-driven Typography, Prosody-driven Typography)
- A design technique in which the visual properties of text — typically font weight, width, or size on a variable-font axis — are modulated in real time by features extracted from a corresponding speech…
- Stenographer (Stenocaptioner, Court Reporter)
- A trained professional who produces real-time verbatim transcription of speech, typically using a stenotype machine that maps chorded key combinations to phonetic syllables. In accessibility contexts,…