TranscriptBalance
Open app

Private transcription without uploads: how Whisper runs in your browser

Updated

With TranscriptBalance, the Whisper speech recognition model runs right in your browser, on your own device. Your audio and video files are never uploaded: the only things that come from the internet are the models, downloaded once and then kept in your browser.

How does Whisper run in a browser?

Whisper is an open speech recognition model from OpenAI (MIT license). TranscriptBalance runs it with transformers.js and ONNX Runtime Web, two libraries that let AI models run inside the browser. There are two ways to do that:

The app checks which way fits when it starts. If the graphics card cannot load a model or drops out, the app switches to the processor where possible. Your file is read in the browser, converted to 16 kHz audio, split at pauses into chunks of up to 28 seconds and transcribed chunk by chunk. Live mode always runs on the processor with small models, which keeps the graphics card free for the AI.

What is downloaded once?

Nothing when you open the page, only when you start a feature. Then your browser downloads:

Everything is stored in your browser (in the Origin Private File System) and loaded from there next time. The app also asks the browser not to clear these files on its own. As with any download, Hugging Face and jsDelivr see your IP address and which file you fetch, but no content.

What never leaves your device?

Since none of this is stored with us, we cannot recover it either. Export anything that matters to you.

How can I check this myself?

  1. Open TranscriptBalance in Chrome, Edge or Firefox and press F12 (or right-click and choose “Inspect”). Switch to the “Network” tab.
  2. Drop a file and start the transcription.
  3. You will see downloads from huggingface.co or hf.co and cdn.jsdelivr.net (mostly the first time), parts of the app from our own domain and short messages to /api/e. There is no request that uploads your file.

Click on an /api/e message to see what it contains: the type of action and rough details such as the model and the length rounded to minutes, never any text or file names.

What leaves your device with optional features?

All the details are in the privacy policy.

Requirements and limits

Transcribe in your browser now

Frequently asked questions

Is in-browser transcription GDPR-compliant?

TranscriptBalance never receives your content: audio, transcripts and file names stay on your device, and no server processes them. If you use a recording for more than private purposes, you are responsible for the data protection of the people in it. For interviews or meetings, get their consent and follow the rules of your university or employer.

Can Hugging Face see what I transcribe?

No. Hugging Face only serves the model files and, as with any download, sees your IP address and which model you fetch. For the German live model, that reveals that you are transcribing German. Audio and text never go to Hugging Face.

Why is the first run slower?

The first time, your browser downloads the Whisper model, between 44 and 762 MB depending on the model. After that it is stored in your browser and TranscriptBalance starts without downloading it again.

How do I delete the models and my transcripts?

In TranscriptBalance’s model picker, delete transcription models one by one with “Delete model”; “Delete all” removes every stored model file. Remove transcripts one by one or with “New”, and saved sessions with “Delete session”. To remove everything at once, clear TranscriptBalance’s site data in your browser.

Why does the “Precise” model need a graphics card?

“Precise” is Whisper large-v3-turbo, the largest of TranscriptBalance’s four models. On the processor it would be too slow for long recordings, so the app only offers it with WebGPU. Without WebGPU, “Balanced” (Whisper small) is the most accurate choice.

Does the AI summary run locally too?

In file mode, yes. TranscriptBalance first uses Chrome’s built-in AI (Gemini Nano) and otherwise Gemma 3 1B via WebGPU, both on your device. Only in live mode can you choose to connect Gemini with your own Google key; excerpts of the transcript then go to Google.