Audio to text uses speech recognition to produce editable text from a file and, where supported by the interface, live microphone capture. It creates a first draft for meetings, interviews, lessons, and voice notes but cannot guarantee verbatim accuracy or identify every speaker.
Accents, overlapping speakers, music, specialist terms, and poor microphones reduce accuracy. Proofread against the audio, especially names, numbers, negations, and accountable statements, and manually confirm timing and speakers in important records.