Ermine.ai
Ermine.ai transcribes audio from your device microphone entirely in the browser with no server upload. You click to start, speak, and get a live transcript plus downloadable audio and text files when you finish. The transcription model loads client-side on first use (about 50 MB download) and caches locally so later sessions start faster.
Cloud transcription services send your audio to remote servers for processing. Ermine.ai runs the model on your machine via transformers.js, which means recordings never leave your device. That privacy trade-off comes with limits: English-only transcription and a one-time model download that can take a few minutes on first load.
Journalists, students, and privacy-conscious professionals use Ermine.ai when they need quick voice notes without signing up for an account or sending audio to a third party. It fits anyone who wants a free, no-login dictation page that works offline after the model caches.
100% client-side transcription with no audio sent to external servers
Download both the recorded audio file and transcript when finished
Transcription model caches locally after first load (~50 MB download)
Runs in the browser with no account or signup required
Open-source project available on GitHub (vishnumenon/ermine-ai)
Audio never leaves your device, which is rare among browser transcription tools
No account, signup, or payment required to start transcribing
Open-source codebase on GitHub for transparency and self-hosting
English-only transcription with no multilingual support
First-time model download (~50 MB) can take several minutes before you can transcribe
Thin feature set compared to cloud services with speaker diarization or editing tools
Is Ermine.ai free to use?
Yes. Ermine.ai is completely free with no subscription or account required. You open the site, load the model, and start transcribing from your microphone at no cost.
Does Ermine.ai upload my audio to a server?
No. Ermine.ai processes all audio locally in your browser using client-side machine learning. Your recordings and transcripts never leave your device during transcription.
What languages does Ermine.ai support?
Ermine.ai supports English transcription only. The site states the loaded model is limited to English, with no other language options available.
How long does Ermine.ai take to load the first time?
The first Ermine.ai session downloads and initializes the transcription model files, which are about 50 MB. This can take a few minutes depending on your connection. Later sessions load much faster from the browser cache.
Can I download my Ermine.ai recording?
Yes. Ermine.ai lets you download both the audio recording and the transcript after you finish a session. Use the Download Audio + Transcript button on the page.

