Live transcription

Transcription turns the interviewer's audio into text you can read while they are still talking.

A rolling caption of the live audio.

How it works

Once a session starts, a rolling caption shows what is being said. Transcription runs in the cloud on every account, so there is no speech engine to install or choose. It helps when you miss a word, and when a question runs long enough that the beginning has already slipped.

Languages

Cloud transcription covers 60+ transcription languages.

Pro's 3,000 credits cover more than 40 live hours on GPT-6 Astra, GPT-6 Luna or Grok 4.7; other models use credits at different rates.

Audio sources

SourceWhat it captures
SpeakerAudio playing on your machine — the interviewer's voice.
MicrophoneYour own voice.

Choose the speaker the interviewer comes through and the microphone you use. Those follow your computer's current devices, and they update when the devices change. Pick a specific one when the call is on a headset or a virtual audio cable.

Mock interviews listen only to your microphone — you are the candidate. Live copilot starts on the interviewer's audio, with your microphone muted. Unmute it when you want your side in the caption too.

Sending a question to the assistant

  • ⌘⌥↵ — send the live caption, or what you typed, plus any attached screenshots.
  • ⌘⌥, — clear the draft text without removing attached screenshots.

Send includes the interviewer’s latest caption, so you do not copy it into the draft first. Attach a screenshot, then send, when the question is on screen.

If detection feels off

Interviewers speak at very different paces. Endpoint timing is set on the server for cloud transcription — there is no pause-interval control in Settings. If captions cut early or run long, check the selected speaker and microphone under audio sources first.

Test audio before the interview
Transcription depends on permissions and on the right device being selected. Run a self-meeting and confirm the caption is moving before you rely on it.