Subtitles and Caption Introduction to Caption 4 — Questions and Answers
Question 1: What is the WebVTT (.vtt) caption format primarily designed for?
- Broadcast television encoding
- HTML5 web video playback (Correct answer)
- DVD menu navigation
- Print transcript generation
Correct answer: HTML5 web video playback
WebVTT (Web Video Text Tracks) was created specifically for use with the HTML5 <track> element to display captions in web browsers.
Question 2: A captioner is unsure whether to caption an aside spoken by an off-screen narrator. Which principle guides this decision?
- Only on-screen speakers should be captioned
- All audio content accessible to a hearing viewer should be captioned (Correct answer)
- Narration is optional and only captioned if requested by the client
- Off-screen audio is described in brackets, not captioned
Correct answer: All audio content accessible to a hearing viewer should be captioned
Captioning best practice requires that any audio a hearing viewer can perceive be made accessible through captions, including off-screen narration.
Question 3: How are non-speech sounds such as a doorbell or explosion typically formatted in professional captions?
- They are omitted unless the director requests them
- They are written in all lowercase inside round parentheses
- They are written in uppercase or title case inside square brackets (Correct answer)
- They are inserted as emoji symbols
Correct answer: They are written in uppercase or title case inside square brackets
Non-speech sounds are conventionally written in square brackets using uppercase or title case, such as [DOORBELL RINGS] or [Explosion], to distinguish them from dialogue.
Question 4: What is the key difference between SDH (Subtitles for the Deaf and Hard of Hearing) and standard subtitles?
- SDH uses a different font; standard subtitles use a default system font
- SDH includes non-speech audio descriptions and speaker IDs; standard subtitles only translate dialogue (Correct answer)
- SDH is delivered via line 21 data; standard subtitles use a separate track
- SDH is only available on Blu-ray; standard subtitles work on all formats
Correct answer: SDH includes non-speech audio descriptions and speaker IDs; standard subtitles only translate dialogue
SDH adds sound effect descriptions and speaker identification labels on top of translated dialogue, making it accessible to deaf viewers unlike standard subtitle tracks.
Question 5: Line 21 of the vertical blanking interval (VBI) is historically significant in captioning because:
- It stores color information for caption text rendering
- It was the analog broadcast channel used to carry closed caption data in NTSC video (Correct answer)
- It defines the frame rate for caption synchronization
- It is the default line where pop-on captions appear on screen
Correct answer: It was the analog broadcast channel used to carry closed caption data in NTSC video
Line 21 of the NTSC VBI carried encoded caption data that TV decoder chips would extract and display, forming the backbone of analog closed captioning.
Question 6: When a speaker's words are inaudible or unintelligible on the audio track, how should the captioner handle it?
- Leave a blank gap with no caption
- Write [INAUDIBLE] or [UNINTELLIGIBLE] in brackets (Correct answer)
- Insert a best guess in regular text format
- Use an ellipsis (…) to indicate missing content
Correct answer: Write [INAUDIBLE] or [UNINTELLIGIBLE] in brackets
Using [INAUDIBLE] or [UNINTELLIGIBLE] in brackets alerts viewers that audio was present but could not be understood, maintaining transparency.
Question 7: Which of the following best describes 'pop-on' captioning?
- Captions that fade in and out gradually
- Complete caption blocks that appear and disappear instantly (Correct answer)
- Captions that scroll from right to left
- Captions that are generated automatically by speech recognition
Correct answer: Complete caption blocks that appear and disappear instantly
Pop-on captions are pre-prepared blocks that appear all at once and then disappear all at once, the standard style for pre-recorded content.
What is the WebVTT (.vtt) caption format primarily designed for?