EarScribe

Audio to text converter

Free audio to text with timestamps

Choose an MP3, WAV, M4A, FLAC, OGG or WebM recording. EarScribe detects the spoken language, creates timestamped text, and lets you search, edit, replay and export the result without uploading the audio.

  • Free and unlimited, with no account or per-minute quota
  • Click any timestamp to check the original recording
  • Copy text or export TXT, SRT, VTT and JSON
Drop an audio file or click to browse
MP3, WAV, M4A, OGG, FLAC, WebM — practical length depends on device memory

Review the transcript before you export

From recording to usable text

1

Choose the recording

Use a browser-supported audio file. On memory-limited devices, shorter files and the balanced model are more reliable.

2

Use the recommended quality

Most users can keep the balanced setting. Advanced model choices stay hidden until you need them.

3

Review the result

Search the transcript, edit names or technical terms, replay uncertain lines, then export the format your next tool needs.

EarScribe timestamped transcript workspace with waveform, search, editing and export controls

What improves transcription quality

  • Clear speech and lower background noise usually matter more than choosing the largest model.

  • For interviews, keep speakers close to the microphone and avoid music underneath speech.

  • Check names, numbers and specialist terms against the recording before publishing.

Questions about this workflow

Which audio formats can I convert to text?

EarScribe accepts common browser-playable formats including MP3, WAV, M4A, OGG, FLAC and WebM.

Do I need to choose the spoken language?

No. Whisper detects the language from the recording. You can verify the detected language in the result workspace.

Is the transcript editable?

Yes. Edit individual timestamped sections, search the text, replay the matching audio and export the edited version.

Related audio tools