<aside> 📌

How to read this doc: These are the default guidelines for all Spoken Form transcriptions unless otherwise specified. 🔑 CORE RULE — the golden principle behind the whole doc 🆕 KEY DIFFERENCE — where Spoken Form differs from the standard/official-format approach

</aside>

General Information

What is Verbatim AI Training — Spoken Form?

Verbatim AI Training (Spoken Form) is a customer-specific service level of transcription. It differs from HappyScribe's standard verbatim guidelines and from the official/standard-format variant: here, numbers, codes, and personal data are transcribed exactly as they are voiced — as words, never converted to digits, symbols, or conventional written formats.

<aside> 🧑‍🤝‍🧑

Quality tip! These guidelines are an addition to the general guidelines for transcription formatting, sentence structure, symbols, italics, etc.

However, the instructions below must always be followed when working on these files unless otherwise specified by Admin.

</aside>

<aside> 🔑

CORE PHILOSOPHY — Preserve Everything. The goal is to retain as much of the target speaker's content as possible. Downstream the customer can always remove disfluencies, fillers, and hesitations — but they cannot recover information that was never transcribed. When in doubt, keep it.

</aside>

<aside> 🆕

KEY DIFFERENCE — Spoken Form vs. Standard Form. Write exactly what was voiced. Transcribe numbers, symbols, and punctuation as the words the speaker actually said, never as digits, symbols, or conventional written formats. Two speakers reading the same item differently produce two different transcriptions.

</aside>

Languages and services

Transcriptions will be required in the following languages:

Speaker identification

To make sure speakers can be distinguished while remaining anonymous:

Use Speaker 1, Speaker 2, etc. as labels. Do not use their name.

If it's a company entity or celebrity, you can still write their name.

Transcription Logic

Complete words

If it is a complete word, transcribe it.

Partial words / False starts

A partial word that is not repeated — the speaker cuts off a word and moves on. Transcribe the audible fragment using two hyphens (--) to mark the cutoff.