VideoText workflow guide

GoTranscript Inaudible Tags: [inaudible], [unintelligible] & [crosstalk]

Which tag when audio is unclear? GoTranscript uses [inaudible 00:00:00] and [unintelligible 00:00:00] — not interchangeable — plus [crosstalk] for overlapping speech. Full rules with examples.

Three tags for three different situations

  • GoTranscript splits unclear audio across two tags where most services use one. [inaudible 00:00:00] covers speech that cannot be heard — noise, a poor recording, a dropped line. [unintelligible 00:00:00] covers speech that comes through clearly but cannot be made out, usually accent or delivery.
  • Overlapping speech is neither of those. [crosstalk] is its own tag, and reaching for [inaudible] because two people talking at once could not be separated is a distinct QA failure.
  • Both unclear-audio tags carry the full hours, minutes and seconds, and both are bold. [inaudible 3:45] has the right tag with the wrong time format, and is rejected on the format alone.

From raw transcript to client-ready formatted file

1. Select and configure the target style guide

Choose Rev, GoTranscript, TranscribeMe, Scribie, or a custom client specification. Each guideline has distinct rules for verbatim level, speaker label format, timestamp intervals, inaudible notation, and maximum paragraph length.

2. Set verbatim level explicitly

Clean verbatim removes fillers, false starts, and repetitions for readability. Full verbatim preserves all spoken content. Applying the wrong level is the most common single reason transcripts fail marketplace QA — it cannot be fixed without re-reading the source audio.

3. Normalize speaker labels throughout the file

A single speaker must have exactly one label format from the first occurrence to the last. Mixed formats (JOHN SMITH / John / J. Smith) require a find-and-replace pass across the full document before any other formatting work.

4. Apply timestamp rules and paragraph breaks

Rev-style: timestamps every 2 minutes or at each speaker change. GoTranscript: job-dependent — bold [00:00:00] either every 2 minutes or at every speaker change, with no timestamping on the qualification test. TranscribeMe: per-speaker-turn timestamps. Paragraph limits range from 8 lines (Rev) to roughly 500 symbols (GoTranscript).

5. Run pre-delivery QA check

Verify inaudible notation consistency (all [inaudible] or all [INAUDIBLE], never mixed), check bracket format for crosstalk sections, confirm verbatim level is consistent throughout, and validate that paragraph length does not exceed the client's maximum.

Tag reference

[inaudible 00:00:00]

Speech that cannot be heard at all. Place it at the exact point in the text where the audio fails rather than at the end of the sentence, so the editor can find the moment in the recording.

[unintelligible 00:00:00]

Speech that is audible but cannot be understood. Choosing this over [inaudible] tells the editor the recording itself was fine and the delivery was not — which is information they act on differently.

[crosstalk]

Two or more speakers talking over each other. No timestamp on this one. Place it inline or on its own line at the point where the overlap happens.

Pauses and silence

[silence] marks a gap of roughly 4 to 10 seconds. Anything longer takes [pause 00:00:00], bold and on its own line, with the time at which the pause begins.

Transcriptionists and editors running formatting workflows

Freelance transcriptionists

Reduce revision risk before submitting marketplace jobs. A pre-delivery formatting check catches the label inconsistency, timestamp format error, or verbatim level mismatch that triggers a QA rejection and unpaid revision.

Agency QA leads and editors

Create a consistent formatting baseline across a team of transcriptionists working on the same client account — so reviewers spend time on content accuracy, not fixing label formats.

Researchers and journalists

Turn raw AI-generated transcripts into readable interview documents: speaker structure, consistent punctuation, accurate paragraph breaks, and timestamps that let readers verify context in the source recording.

Tag mistakes editors look for

Guessing the word

A plausible word typed in where the audio could not be heard. This is worse than any tag: an editor checking the file cannot tell that it is wrong, so it passes review and reaches the client as an error.

Short-form time

[inaudible 3:45] rather than [inaudible 00:03:45]. The full hours, minutes and seconds are required on every time-stamped tag, on short files as much as long ones.

One tag doing every job

Every unclear moment marked [inaudible], including audible-but-unclear speech and passages where speakers overlap. It collapses three distinctions the guidelines deliberately keep apart.

Invented notation

[unclear], [mumbles], [??] and similar improvisations. The guidelines list the tags that exist, and anything outside the list is treated as an error rather than a judgement call.

Sound-event notation rules

The shape of a sound note

Square brackets, lowercase, present tense, two words at most: [background noise], [laughs], [laughter], [foreign language]. The constraint is on the notation, not just the vocabulary.

Which notes carry a time

Only the time-stamped marks: [inaudible 00:00:00], [unintelligible 00:00:00] and [pause 00:00:00]. [crosstalk], [silence] and [laughs] take no time and should not be given one.

What bold covers

Bold applies to speaker labels, timestamps and time-stamped tags. A tag that loses its bold on export has lost a graded rule, even though the text still reads correctly.

Transcript formatting and style guide questions

What is the difference between inaudible and unintelligible on GoTranscript?

[inaudible 00:00:00] — speech cannot be heard (noise, poor recording). [unintelligible 00:00:00] — speech is heard but not understood (accent, manner). Both require full HH:MM:SS and must be bold.

How do I mark crosstalk on GoTranscript?

Use [crosstalk] when speakers talk over each other. Place it inline or on its own line where the overlap occurs. Do not use [inaudible] for simultaneous speech — editors treat these as different tags.

Can I use plain [inaudible] without a timestamp?

No. GoTranscript requires [inaudible 00:00:00] with the full hours:minutes:seconds timestamp at the point of unclear audio. Short forms like [inaudible 3:45] are incorrect.

What sound events besides inaudible does GoTranscript allow?

[silence] for 4–10 second pauses, [pause 00:00:00] bold on a separate line for pauses over 10 seconds, [background noise], [laughs], [laughter], and [foreign language] when applicable. Notes are lowercase, present tense, max two words.

Should I guess at unclear words on GoTranscript?

Never guess. Use [inaudible 00:00:00] or [unintelligible 00:00:00] at the exact position. Inventing custom tags or leaving blanks causes QA failures.

Related formatting and QA workflows