Which languages does it understand?
English, Spanish, Portuguese, French, German, Italian, Dutch, Polish, Romanian, Czech, Swedish, Russian, Ukrainian and Turkish. With Auto-detect it recognises which of them is spoken from the first words.
Turn an interview, a lecture, a voice note or a podcast into text, with the time of every line, in English and 13 other languages. The speech model runs in your browser, so the recording never leaves your device.
English, Spanish, Portuguese, French, German, Italian, Dutch, Polish, Romanian, Czech, Swedish, Russian, Ukrainian and Turkish. With Auto-detect it recognises which of them is spoken from the first words.
No. The page downloads the speech model once (your browser keeps it), and the model listens on your device. The recording and the text stay with you; nobody else hears or reads them.
It depends on your device: on a recent computer, often less than the recording lasts. Long files are cut into pieces of up to half a minute at the pauses, and it works through them one by one.
Clear speech close to the microphone comes out well; music, overlapping voices and strong noise make mistakes more likely. Removing the background noise first helps. Read it through before you use it, as with any automatic transcript.
Yes: Save .srt or Save .vtt gives subtitles with these times. The video editor can also write and time subtitles for a video directly.