MP3 · WAV · M4A · FLAC → TXT · SRT · VTT

Audio to Text

Transcribe interviews, podcasts, lectures and voice memos in your browser. Edit the transcript and export TXT, SRT or VTT.

Select the language spoken in the recording; it is independent of the page language.

First use needs internet to download the Whisper speech model from Hugging Face and the AI runtime from third-party jsDelivr CDNs. Your browser may cache these files. Keep the tab open; processing speed depends on your device.

Your audio and transcript stay on your device and are not uploaded.

First use needs internet to download the Whisper speech model from Hugging Face and the AI runtime from third-party jsDelivr CDNs. Your browser may cache these files. Keep the tab open; processing speed depends on your device.

Review the result: accuracy is not guaranteed. Speaker identification is not supported. This tool does not translate, summarize or separate speakers.

Transcribe an audio recording in three steps

  1. 1

    Choose audio

    Choose an audio recording with clear speech from your device.

  2. 2

    Transcribe audio

    Select the recording’s spoken language and start local transcription.

  3. 3

    Download TXT

    Review and edit the text, then copy or download it.

Ways to use your transcript

  • Turn an interview recording into searchable text.
  • Draft podcast transcripts for editing and accessibility.
  • Review lecture recordings or turn voice memos into notes.

Frequently asked questions

Is my audio recording uploaded?

Your audio and transcript stay on your device and are not uploaded. First use needs internet to download the Whisper speech model from Hugging Face and the AI runtime from third-party jsDelivr CDNs. Your browser may cache these files. Keep the tab open; processing speed depends on your device.

Which audio recordings can I transcribe?

One audio file, up to 100 MiB and 5 minutes. Supported formats: MP3, M4A, WAV, AAC, OGG, OGA, Opus, FLAC, WebM, WebA.

Which language should I choose?

Select the language spoken in the recording; it is independent of the page language.

How accurate is the transcript?

Review the result: accuracy is not guaranteed. Speaker identification is not supported. This tool does not translate, summarize or separate speakers.

Select Region

Choose your preferred language and region

Audio to Text – TXT, SRT, VTT | videotolink