v2.0Choose the speech model that suits your PC

Every meeting, in English.

Dolmi listens to Teams, Zoom, Meet or any app on your PC and shows live English subtitles in a floating bar. Speech recognition and translation run on your PC, so your audio never leaves it.

v2.0.0 · 97 MB · Windows 10 / 11 · free and open source (MIT)

Dolmi's live view: German speech on the left, the English translation on the right, with line count, word count and lag.
The floating subtitle bar showing an English line with the German original underneath.
THE FLOATING SUBTITLE BAR ↑
27+spoken languages, German by default, or auto-detect
On your PCspeech recognition and translation, offline once models are in
No accountno sign-up, no telemetry, no analytics
MITopen source, built and install-tested in public CI
Features

Built for the meeting you're in now.

Subtitles while people talk, a record you can trust afterwards, and an assistant if you want one.

Meetings

Every meeting saved in both languages.

Each session is kept as Markdown and SRT, in the original language and in English, in your Documents folder.

  • Copy, archive, delete or summarize any meeting
  • Turn saving off, or auto-delete after 7, 30 or 90 days
Dolmi's Meetings view listing saved sessions with copy, archive, delete and summarize actions.
Assistant · optional

Ask about the meeting while it happens.

"What did Jonas commit to?" The assistant answers from the live transcript, using your own Claude or OpenAI key. Without a key, Dolmi sends nothing anywhere.

  • Chats are searchable and can be copied or archived
  • Only English text is sent, never audio
Dolmi's Assistant answering a question about the current meeting.
This PC & speech models

Dolmi checks your PC before it downloads anything.

It reads your GPU, RAM and CPU, says how well each Whisper model will run on this machine, and asks before any large download.

  • Download, switch or remove models in one click
  • NVIDIA GPU? Large v3 Turbo runs live
Dolmi's model settings showing detected hardware and how well each speech model fits this PC.

Any meeting app

Dolmi hears what Windows plays, so it works with Teams, Zoom, Meet, Slack or a browser tab. You still hear everything, and you don't need a virtual audio cable.

Off screen shares

Invisible mode keeps the window and subtitle bar out of screen shares, recordings and the taskbar. Ctrl+Alt+Shift+D turns it off.

Vocabulary & glossary

Teach Dolmi your names, products and terms so they're spelled and translated correctly. Changes apply without a restart.

Clean up automatically

Turn off saving entirely, or have old transcripts deleted after 7, 30 or 90 days.

Paper and Ink

A light and a dark theme, an adjustable subtitle size and opacity, and a bar you can drag and resize anywhere.

A proper Windows app

An installer with a Start-menu shortcut, clean upgrades and uninstall, and no admin rights needed.

How it works

From speaker to subtitle.

Every step runs on your PC. No PyTorch, no cloud speech service.

Capture

Records what Windows plays. You still hear it.

WASAPI loopback

Split

Cuts the audio at natural pauses.

Silero VAD

Transcribe

Turns speech into text in the spoken language.

faster-whisper

Translate

Translates each line into English.

Opus-MT · CTranslate2

Show & save

The subtitle bar, plus a transcript in both languages.

.md + .srt
Privacy

Your audio stays home.

Dolmi uses the network only for things you start yourself: downloading a model, or the optional assistant.

Local speech

Recognition and translation run on your PC. Audio is never uploaded.

No account

No sign-up, no telemetry, no analytics, and no tracking of any kind.

AI is optional

The assistant uses your own key and sends English text only. Without a key, nothing is sent.

Your files

Transcripts live in your Documents folder. Delete them yourself, or let Dolmi do it on a schedule.

Transcribing other people processes their personal data, so tell meeting participants you're using live transcription. Dolmi reminds you once. Read the full privacy policy.

Speech models

Pick for your PC, or let Dolmi pick.

Models download once and then work offline. Each language also needs a translator of about 150 MB.

ModelSizeGood for
Whisper Small AUTO · CPU486 MBPCs without an NVIDIA GPU
Whisper Medium1.5 GBBetter names and terms; adds 2–4 s on a CPU
Whisper Large v3 Turbo AUTO · GPU1.6 GBNVIDIA GPU: near-best accuracy, still fast
Whisper Large v33.1 GBHighest accuracy; needs an NVIDIA GPU for live use
Install

Running in two minutes.

  1. Download the installerDolmi-Setup-2.0.0.exe from GitHub Releases.
  2. Run it, no admin neededIt installs for your user. Tick GPU speech recognition if you have an NVIDIA card (about 1.3 GB).
  3. Press StartDolmi suggests the model that suits your PC and asks before downloading it.
System
Windows 10 (2004+) or 11, 64-bit
Memory
8 GB RAM, 16 GB recommended
GPU
Not needed. NVIDIA with 4 GB+ VRAM for Large v3 Turbo
Audio
Any speakers or headphones

Windows may warn you. The installer isn't code-signed yet, so you may see "Windows protected your PC". Click More info → Run anyway. Every build comes from a tagged commit and is install-tested in public CI.

FAQ

Questions, answered.

Does Dolmi need the internet?

Only to download speech and translation models the first time, and for the optional assistant. After that, live subtitles work fully offline.

Which languages does it understand?

27 spoken languages plus auto-detect, with German as the default. Subtitles are always in English. A few languages without their own translator use a multilingual model, which makes more mistakes.

Do I need a graphics card?

No. Whisper Small runs well on a normal CPU. An NVIDIA GPU with 4 GB or more lets you use Large v3 Turbo, which is more accurate and still fast.

Will other people in the call see the subtitles?

Not if you turn on invisible mode, which leaves Dolmi out of screen shares and recordings. Please still tell people you're using live transcription.

Is it really free?

Yes. Dolmi is MIT-licensed and the source is on GitHub. The optional assistant uses your own AI provider key, so any cost there is between you and the provider.

Free for Windows

Understand every meeting.