A free subtitle generator with no cap and no watermark

“Free” does a lot of work in this market. Usually it means free for the first few minutes, or free with branding burned into your video, or free once you hand over an email address. This page explains what free means here, why the economics are different, and — since every tool has a catch — exactly what you give up in exchange.

No sign-up step to get past

Open the page, drop in a video, download the subtitles.

Generate subtitles free →

Why hosted tools have to put limits somewhere

This is not a complaint about other services — it is arithmetic. Running speech recognition on a server costs the operator real money: GPU time, bandwidth to receive your video, storage while it is processed. Those costs scale directly with minutes of audio transcribed. A service that gave away unlimited transcription would be paying, per user, without limit.

So the cost gets recovered, and there are only so many ways to do it:

What changes when the model runs on your machine

The entire cost structure above disappears, because none of it is being paid. There is no server receiving your video — the page is static HTML, and the Whisper model is downloaded into your browser and executed there, on your own CPU or GPU. Transcribing ten hours of audio costs the site exactly as much as transcribing ten seconds: nothing.

So the honest claim list is short and checkable:

 Here
Minute or length capNone. The limit is your hardware and your patience.
Watermark on exportNone, on the video or in the subtitle file.
Account / emailNot required. There is nothing to sign up to.
Payment or card detailsNever requested. There is no paid tier.
Export paywallNone. SRT, VTT and burned-in MP4 are all free.
Your videoRead from disk by the page. Not uploaded, because there is no server to upload to.
Tracking / analyticsNone on the page.

That last row is worth a note: the page does fetch two things from third-party CDNs — the Whisper model weights, and the WebAssembly builds of the inference runtime and ffmpeg. Those are downloads to you. Nothing about your video, your audio or your subtitles is sent anywhere. How to verify that yourself in about thirty seconds.

What you give up

Every tool has a catch, and pretending otherwise would be the sort of claim this page is complaining about. The catch here is that you supply the compute, and that has four concrete consequences.

  1. A model download, once. 75 MB, 145 MB or 480 MB depending on which size you pick. Your browser caches it, so it is a one-time cost per model — but it is a real wait the first time, and a real amount of data on a metered connection.
  2. Speed is your hardware’s speed. A recent laptop with WebGPU is quick. An old machine running on the CPU is not. A hosted service with a rack of GPUs will beat your laptop on wall-clock time, and there is no way around that.
  3. Smaller models than a paid API runs. Browser-practical Whisper tops out at small; a paid API runs large. On clear speech the gap is modest. On accented, noisy or overlapping speech it is noticeable. The comparison, with numbers.
  4. You will edit the output. True of every automatic subtitle tool at any price, but worth setting the expectation: proper nouns, acronyms and jargon need a pass, and stretches of silence sometimes produce invented lines that need deleting.

There is also a hard requirement rather than a trade-off: it needs a reasonably current browser with WebAssembly and Web Workers, and it will not run from a file:// URL.

When a paid service is the right answer

Worth saying plainly. If you are captioning a large volume of difficult audio to a professional standard, on a deadline, a paid service running a large model with human review is the correct tool and the money is well spent. This is not that. This is the right tool when you have a video, you want decent subtitles, you do not want to upload the file, and you do not want to pay or sign up for either.

Related