Creating and editing captions
Caption metadata
The ability to add caption metadata allows you to add information to your captions beyond what is being said at that point in time. CaptionHub currently supports two types of metadata:
- Speaker identification
- New paragraph flags
Speaker detection
If you've enabled Speaker detection, our speech recognition engine will attempt to identify who's speaking and assign a label to their captions.
CaptionHub labels the first speaker as S1, followed by S2, S3 and so on. These labels provide a useful starting point, but you can update them to reflect the speaker's name.
To assign or edit a speaker:
- Click the speaker icon next to the caption number in the spreadsheet view on the Edit page. This opens a pop-up window. The icon will appear differently depending on whether a speaker has already been assigned:
- No speaker assigned
- Speaker assigned


- Select a speaker from the existing list, or click Add new speaker to create one.

- Alternatively, enter the speaker's name directly in the text field and click Add and assign.


- To edit an existing speaker's name, click the pencil icon.

- To remove a speaker from a caption, click Remove.
- To assign a different speaker, click the speaker icon and select another speaker from the list.
Speakers are colour-coded to make them easier to identify throughout the caption set.
Note: The current transcript export format that supports and reflects changes made to speaker identification is DOCX (Word).
New paragraph flags
Transcribed text can be difficult to read when presented as a single block of text. For supported export formats, you can use paragraph flags to indicate where a new paragraph should begin.
To add a new paragraph, click the paragraph button. This marks the caption as the start of a new paragraph in your transcript.
To remove the paragraph flag, click the button again.

