Choose the recording
Use a browser-supported audio file. On memory-limited devices, shorter files and the balanced model are more reliable.
Audio to text converter
Choose an MP3, WAV, M4A, FLAC, OGG or WebM recording. EarScribe detects the spoken language, creates timestamped text, and lets you search, edit, replay and export the result without uploading the audio.
Review the transcript before you export
Use a browser-supported audio file. On memory-limited devices, shorter files and the balanced model are more reliable.
Most users can keep the balanced setting. Advanced model choices stay hidden until you need them.
Search the transcript, edit names or technical terms, replay uncertain lines, then export the format your next tool needs.

Clear speech and lower background noise usually matter more than choosing the largest model.
For interviews, keep speakers close to the microphone and avoid music underneath speech.
Check names, numbers and specialist terms against the recording before publishing.
EarScribe accepts common browser-playable formats including MP3, WAV, M4A, OGG, FLAC and WebM.
No. Whisper detects the language from the recording. You can verify the detected language in the result workspace.
Yes. Edit individual timestamped sections, search the text, replay the matching audio and export the edited version.