How It Works

From Recording To Finished Transcript In Four Steps

No setup, no plugins, no waiting for a sales call. Create an account, add your media, and read a speaker-labelled transcript with timestamps you can click.

Free forever plan — no credit card required.

The Workflow

Four Steps, Start To Finish

Each step is designed so you can walk away — long jobs continue in the background and finish without you watching.

  1. Bring in your media

    01

    Drag and drop audio or video, paste a link to a video or podcast, pick files from your device, or queue a whole batch at once. Video is converted to audio in your browser first, so a multi-gigabyte file uploads in a fraction of the time.

  2. Choose your options

    02

    Set the spoken language, speaker count, speaker sensitivity and roles, and switch on filler removal. Key terms, PII redaction and multi-channel need Pro or Business. In bulk, set one default and override individual files.

  3. We transcribe it

    03

    Your recording runs through a state-of-the-art speech engine with speaker separation and word-level timestamps. Paid plans get priority in the queue, and jobs keep running if you close the tab.

  4. Read, refine, and reuse

    04

    Open the transcript with synced playback, rename speakers, generate a summary and action items, translate it, then export in the format your workflow needs.

Getting Media In

Three Ways To Start A Transcript

Whatever the source, the same validation, progress feedback and options apply.

01

Upload files

MP3, M4A, WAV, FLAC, OGG, MP4, MOV, MKV, WEBM and more, up to 3 GB per file on Free and 5 GB on Pro and Business. Progress, cancel and retry are all visible per file.

03

Bulk queue

Queue dozens of files with per-file language and shared options. Each item shows its own progress and can be retried or removed without disturbing the rest.

After Transcription

What You Can Do With The Transcript

The transcript is the beginning of the work, not the end of it.

Search and jump

Full-text search across the transcript with click-to-play on any sentence, thanks to word-level timestamps.

AI summaries

Summaries, chapters, action items and Q&A on Pro and Business, generated from the finished transcript.

Translation

On Pro and Business, translate a transcript into 100+ languages, or translate the original audio into English while it is transcribed.

All 15 export formats

15 formats — TXT, Word, PDF, SRT, VTT, YouTube and broadcast captions, Markdown, HTML, CSV, Excel, JSON and speaker-grouped text — on every plan, including Free.

Team sharing

On Business, share individual transcripts or whole folders with one or more workspaces, or create a public read-only link, with an activity log of who did what.

Structured data

JSON and CSV exports keep speakers, timings and confidence so you can pipe transcripts into your own tools.

Accuracy Tips

Small Choices That Make A Big Difference

The engine does most of the work, but three habits reliably improve the result.

  • Add key terms before you start

    Names, acronyms and product vocabulary are the words a general model is least likely to guess correctly.

  • Set the language explicitly

    Detection is good, but a bilingual opening minute is exactly where it can commit to the wrong choice.

  • Tell it how many speakers there are

    For panels and interviews, an expected speaker count and roles keep labels stable instead of drifting mid-recording.

Built For Real Volume

Unlimited transcription time on every plan, folders to keep projects apart, bulk sharing for teams, and a background worker that keeps queued jobs moving after you close the browser.

Still have questions? Read the FAQ.

Try It On Your Hardest Recording

The Free plan includes speaker separation, timestamps and all 15 export formats. No card required.

Transcribe Free