Live Captions
Live captions are text versions of spoken words that appear on screen in real time as someone talks. They are commonly used during live events, lectures, meetings, and video calls to help people who are deaf or hard of hearing, as well as others, follow the audio. Many devices and platforms now include a built-in live captions feature that automatically converts speech to text.
Live captions refer to the real-time conversion of spoken audio into synchronized on-screen text for live presentations, lectures, meetings, and streamed content. Implementations range from operating-system and application-level features that provide automatic speech-to-text transcription (for example, device-level live captions built into iPhone/iOS and Windows, and browser extensions using the device microphone) to professionally produced live captioning workflows for lectures and events. Automatic (machine-generated) live captions vary in accuracy and may not meet the correctness or completeness expectations of human-produced real-time captions; practitioners should evaluate the method used against the intended use case rather than assuming any single approach satisfies a given accessibility requirement. This entry describes the feature category generally and is not legal advice.
Why it matters
Live captions make spoken audio accessible in real time, which is essential for people who are deaf or hard of hearing to follow live lectures, meetings, events, and streamed content as they happen. Because these settings are inherently time-sensitive, captions that appear synchronized with the audio allow participants to engage without waiting for a transcript to be produced afterward. Live captions also assist others, such as people in noisy or sound-sensitive environments and those who process written text more easily.
The method used to generate live captions matters significantly. Automatic (machine-generated) live captions vary in accuracy and may not meet the correctness or completeness expectations of human-produced real-time captioning, particularly where specialized vocabulary, multiple speakers, or poor audio quality are involved. Relying on automatic captions in high-stakes settings, such as academic lectures or public events, may leave gaps that affect a person's ability to fully understand the content.
Because of this variability, practitioners should evaluate the captioning approach against the intended use case rather than assuming any single implementation satisfies a given accessibility need or requirement. Whether a particular method meets applicable legal or institutional obligations depends on the specific context, jurisdiction, and current guidance; this entry describes the feature category generally and is not legal advice. Organizations with questions about their obligations should consult qualified legal counsel.
Who it's relevant to
Inside Live Captions
Common questions
Answers to the questions practitioners most commonly ask about Live Captions.