99 languages · Unlimited on paid plans

Turn speech into text, summaries, and subtitles in minutes

Upload audio or video and get an accurate transcript with speakers separated, a summary with the actual decisions, and subtitles ready to ship. No per-feature upcharges.

See pricing
  • No credit card to start
  • GDPR · EU hosting
  • Never trained on your data

quarterly-review.mp4

42:18 · 3 speakers · done in 4:12

99% · excellent

00:14MayaRevenue grew fifteen percent, and the retention metrics are green.

00:31DanielGood. What are we committing to for Q2?

00:38PriyaConcrete steps by Friday. I'll own the rollout plan.

AI summary · action items

  • Priya — rollout plan by Friday
  • Daniel — approve Q2 budget

How it works

Three steps. None of them yours.

  1. 01

    Drop the file in

    Audio or video, any common format, up to 5 GB. A link to a recording works too.

  2. 02

    The engine does the work

    Language detection, transcription, speaker separation, punctuation, and timecodes. Nothing to configure.

  3. 03

    Take what you need

    Clean text, a dialogue split by speaker, an AI summary with decisions, or subtitles — export in a click.

Everything included

The features other tools sell as add-ons are just part of the plan.

Diarization, summaries, and subtitles are often billed separately elsewhere. Here they are included on every paid plan, from the first minute.

  • Speaker diarization

    Up to 6 speakers separated automatically, with labels you can rename.

  • AI summaries

    Decisions, action items, and owners — not a paraphrase of the whole call.

  • Subtitles

    SRT, VTT, and ASS, or burned straight into the MP4 without a watermark.

  • Word-level timecodes

    Every word carries its own timestamp, so editing stays in sync with audio.

  • 99 languages

    Automatic detection, including mixed-language recordings.

  • Export anywhere

    Word, PDF, TXT, and subtitle formats. Your transcript is not locked in.

What we can prove

The numbers we'll stand behind.

NUMU PRO is new, so there are no customer quotes here yet — we'd rather show the engine's actual characteristics than invent testimonials.

99%accuracy on clean audio
Measured on clear recordings. Noisy rooms and heavy accents land lower — we say so rather than promise a number we can't hold.
99languages supported
Detected automatically, including recordings that switch language mid-sentence.
6speakers separated
Diarization labels each voice so a meeting reads as a dialogue, not a wall of text.
EUhosting and data residency
GDPR compliant, DPA available, and your recordings are never used to train models.

Pricing

Clear pricing. Free to start.

Every paid plan includes unlimited transcription and every feature. What changes is team size, API access, and support.

Billing period
  • Free

    Try the engine on your own files

    $0/mo$0/mo

    Free forever

    Volume
    60 minutes / month
    Seats
    1 user
    Max file
    30 MB
  • Starter

    Unlimited transcription for one person

    $8.50/mo$4.25/mo

    Billed monthly$51 billed yearly

    Volume
    Unlimited*
    Seats
    1 user
    Max file
    1 GB
  • Business

    Team workspace with SSO and roles

    $59/mo$38.94/mo

    Billed monthly$467.28 billed yearly

    Volume
    Unlimited*
    Seats
    5 seats
    Max file
    5 GB
  • Enterprise

    On-premise, air-gapped, or custom volume

    Custom

    Priced to your volume

    Volume
    Unlimited
    Seats
    Custom
    Max file
    Custom

* Unlimited transcription on all paid plans, subject to a fair-use clause in the terms of service.

FAQ

Common questions

How accurate is the transcription?

Around 99% on clear recordings — a good microphone, one person speaking at a time. Background noise, strong accents, and heavy crosstalk lower that, as they do for every engine. You get 60 free minutes precisely so you can measure it on your own audio instead of taking our word for it.

What does "unlimited" actually mean?

Unlimited transcription minutes on every paid plan. There is a fair-use clause in the terms to stop abuse at extreme volume — it exists for resellers pushing tens of thousands of hours, not for normal work. If your usage is ever a problem, we contact you before doing anything.

Which languages are supported?

99 languages, detected automatically. Recordings that switch language mid-sentence are handled too, which matters for multilingual teams.

Do you train models on my recordings?

No. Your audio and transcripts are never used to train models, on any plan including Free. Data is hosted in the EU, we are GDPR compliant, and a DPA is available.

Can I get subtitles, not just text?

Yes. SRT, VTT, and ASS files on every paid plan, or subtitles burned directly into the MP4 without a watermark. The Free plan exports SRT and adds a watermark on burn-in.

Is there an API?

Yes, from the Pro plan up. Send a file, get JSON back with text, speakers, word-level timings, and a summary. Diarization and subtitles are included in the per-minute rate rather than billed as extras.

What happens when I hit my Free limit?

Nothing breaks and nothing is deleted. You keep access to everything already transcribed, and you can either wait for the monthly reset, buy minutes as you go from $0.10, or move to a paid plan.

Upload your first file now.

Sixty minutes free every month, no card required. See what the transcript looks like on your own audio before you decide anything.

Compare plans

Not live yet

The service is launching soon

Accounts aren't open yet — we're finishing the last round of testing. Leave us a note and we'll email you the moment sign-ups open.

Email me at launch