Skip to content
Speech to Text

Speech to text, live

Say it once. Read it back as it happens.

Stream straight from your microphone over a live connection, or drop in a recording — Speech to Text turns audio into a transcript in real time, using Google's streaming recognition and OpenAI's Whisper.

No install. Sign in with a one-time code and start talking.

From a sentence to a transcript

Four steps, and the only one you do is the first.

  1. Speak, or upload

    Start talking into your microphone, or drop in a .wav recording. Both land on the same screen.

  2. Audio goes out over a live connection

    Microphone audio streams over WebRTC as you talk; an uploaded file is posted once, as a whole.

  3. Google and OpenAI do the listening

    Live audio is streamed to Google Cloud Speech-to-Text; uploaded recordings are transcribed by OpenAI's Whisper model.

  4. Text lands as it's heard

    Interim words firm up into final lines in the transcript, continuously, with no page reload.

Built on the engines doing the actual work

Two ways to get audio in

Talk live over WebRTC, or upload a .wav file. The transcript view is the same either way.

Real engines, not a mockup

Google Cloud Speech-to-Text handles streaming recognition; OpenAI's Whisper handles uploaded recordings.

A proper peer connection

Microphone audio travels over WebRTC, the same transport behind video calls — not a chain of file uploads.

In with a code, not a password

Sign in with a one-time code sent to your email. Nothing to remember, nothing to reset.

Hear it work on your own voice

Sign in, press start, and talk — the transcript fills in as you go.

Open the live demo
Built by WebMob Technologies