Captions
Captions are lines of text displayed on a video or other media that transcribe the spoken words and identify other important sounds, such as music or sound effects. They allow people who are deaf or hard of hearing, as well as viewers in sound-off environments, to follow along with audio content. Captioning is the process of converting a program's audio into this synchronized text.
Captions are a time-synchronized text alternative for the audio track of media content, presenting spoken dialogue along with non-speech audio information (for example, speaker identification, sound effects, and music) needed to understand the content. Captioning is the process of converting audio content from broadcasts, webcasts, films, video, live events, and similar productions into this text. Captions are commonly distinguished as 'closed' (able to be turned on or off by the viewer) or 'open' (permanently rendered into the video). Captions differ from subtitles, which traditionally convey only dialogue translation and assume the viewer can hear other audio. Note that captions are one component of media accessibility and support conformance-related goals for time-based media; a full technical treatment of applicable success criteria and conformance levels is outside the scope of this core definition.
Why it matters
Captions provide access to audio content for people who are deaf or hard of hearing, allowing them to follow spoken dialogue and understand important non-speech sounds such as music, sound effects, and speaker identification. Without captions, video and other time-based media can exclude a significant portion of the audience from information, entertainment, education, and services delivered through audio. Captions also benefit people who are not disabled, including viewers in sound-off environments such as public spaces, offices, or transit, and those watching content in a language they are still learning.
In the context of digital accessibility, captions are one of the primary components of accessible time-based media and are commonly referenced when organizations work toward WCAG-related goals for video content. Providing captions supports the goal of offering a text alternative to audio so that content is perceivable to users who cannot rely on sound. However, captions alone do not address every media accessibility need; for example, they do not convey visual information to users who cannot see the screen, which is handled through other techniques such as audio description.
Because the specific obligations to caption content can depend on the applicable legal framework, jurisdiction, and setting, organizations should treat captioning as part of a broader accessibility strategy rather than a standalone guarantee of compliance. Requirements evolve through regulation and case law, and this entry is not legal advice; readers with specific obligations should consult qualified legal counsel or current agency guidance.
Who it's relevant to
Inside Captions
Common questions
Answers to the questions practitioners most commonly ask about Captions.