Skip to content

Use voice input

Your spoken question is locally transcribed and inserted into the chat input for you to review before sending.

  • Microphone permission for Synthezia.
  • An installed local Whisper model.
  • An open session chat or Global Chat input.
  1. Select the microphone button beside the chat input.
  2. Speak your question. The control shows that recording is active.
  3. Select it again to stop and transcribe your speech.
  4. Review and edit the inserted text.
  5. Send it only after the text accurately reflects your question.

Synthezia records microphone-only voice input, transcribes it locally, and places the result in the chat input. It does not capture system audio for this workflow.

The voice-input transcription is local. Sending the resulting chat question follows the active Local Mode or External API chat boundary.

  • The microphone button is unavailable: wait for an in-progress chat or voice transcription to finish.
  • The input is blank or inaccurate: check microphone permission and use a clearer, shorter utterance before sending.
  • A model dialog appears: download the requested local Whisper model, then retry.