EarScribe

Audio to text converter

Free audio to text with timestamps

Choose an MP3, WAV, M4A, FLAC, OGG or WebM recording. EarScribe detects the spoken language, creates timestamped text, and lets you search, edit, replay and export the result without uploading the audio.

  • Free and unlimited, with no account or per-minute quota
  • Click any timestamp to check the original recording
  • Copy text or export TXT, SRT, VTT and JSON
Drop an audio file or click to browse
MP3, WAV, M4A, OGG, FLAC, WebM — practical length depends on device memory

Review the transcript before you export

From recording to usable text

1

Choose the recording

Use a browser-supported audio file. On memory-limited devices, shorter files and the balanced model are more reliable.

2

Use the recommended quality

Most users can keep the balanced setting. Advanced model choices stay hidden until you need them.

3

Review the result

Search the transcript, edit names or technical terms, replay uncertain lines, then export the format your next tool needs.

EarScribe timestamped transcript workspace with waveform, search, editing and export controls

audio to text with timestamps

Make an audio transcript you can check

A useful transcript is more than a paragraph of guessed words. It gives you a way back to the recording, so you can verify a quote, fix a name, or find the exact moment before you share it.

Start with the recording you actually have

MP3, WAV, M4A, FLAC, OGG and WebM files can be selected directly. Clear speech and a microphone close to the speaker usually improve the result more than a larger model does.

The first run downloads the selected Whisper model. After that, the browser may reuse its cache. Long files still depend on available device memory, so trimming long silence is a practical first step.

Use timestamps as an evidence trail

Search the transcript for a person, number or phrase, then click its timestamp to replay the matching audio. This is the fastest way to catch a plausible-looking mistake that spellcheck would miss.

For publication, compare names, dates, prices and specialist terms with the source. EarScribe helps you locate the evidence; it does not replace a human review of important claims.

Export the same reviewed text in the right format

Use TXT when you need a clean document, SRT or VTT when the timing must travel with subtitles, and JSON when you need the raw segments and timing for another workflow.

Edits made in the workspace are used by copy and export, so you can correct the transcript once instead of repairing several downloaded files.

Timestamped audio transcript with waveform, search and export controls
Timestamped audio transcript with waveform, search and export controls

Review the transcript before you export

A quick review before you publish

  1. 1

    Names

    Replay every person, company and place name that matters to the story.

  2. 2

    Numbers

    Check dates, prices, measurements and percentages against the audio.

  3. 3

    Context

    Listen to the sentence before and after an excerpt you plan to quote.

  4. 4

    Output

    Choose TXT, SRT, VTT or JSON based on the next tool, not the file extension you started with.

What this page does not promise

  • Local processing reduces upload exposure, but your browser, device, extensions and exported files still need normal security controls.
  • Whisper can mishear overlapping speakers, music, heavy noise and specialist vocabulary. Keep the source recording for review.

What improves transcription quality

  • Clear speech and lower background noise usually matter more than choosing the largest model.

  • For interviews, keep speakers close to the microphone and avoid music underneath speech.

  • Check names, numbers and specialist terms against the recording before publishing.

Questions about this workflow

Which audio formats can I convert to text?

EarScribe accepts common browser-playable formats including MP3, WAV, M4A, OGG, FLAC and WebM.

Do I need to choose the spoken language?

No. Whisper detects the language from the recording. You can verify the detected language in the result workspace.

Is the transcript editable?

Yes. Edit individual timestamped sections, search the text, replay the matching audio and export the edited version.

Related audio tools