Open Speech Studio — tekst-naar-spraak met lokale stemmen
Open Speech Studio — Whisper-spraakmodellen beheren
Open Speech Studio — instellingen met GPU-versnelling

Use cases for AEC professionals

Whether you're inspecting on-site, taking minutes in a project meeting, or dictating an email between tasks - Open Speech Studio works everywhere, even offline.

JOB SITE

On-site inspections

Dictate findings directly into Open Field Studio while taking photos. No taking off gloves, no typing on a tiny screen. Just talk - done.

MEETINGS

Meeting minutes

Real-time transcription of project meetings on location. Fully processed locally - no privacy concerns about client names, prices, or confidential project data in the cloud.

COMMUNICATION

Quick emails on-site

Dictate into Outlook or Gmail between tasks. Hands-free when your hands are dirty, you're wearing safety gloves, or holding a ladder.

AI WORKFLOW

Writing AI prompts

Dictate a prompt for Claude or ChatGPT much faster than typing. Speaking is ~3x faster and you often phrase more naturally - ideal for long, detailed prompts.

DEVELOPERS

Code comments and commits

Dictate code explanations, docstrings, and commit messages directly into your IDE or terminal. Works seamlessly in VS Code, Cursor, Neovim - anywhere you can enter text.

ACCESSIBILITY

Accessibility

For RSI patients, people with motor impairments, or temporary injuries. Work your entire computer completely keyboard-free - GDPR-safe and subscription-free.

Why local?

Cloud speech recognition sends your voice to third-party servers. Open Speech Studio does everything on your own PC - faster, safer, and without monthly costs.

🔒

Privacy

Your voice never leaves your computer. GDPR-safe automatically - no DPA needed with external providers, no risk of data breaches.

💶

No subscription

Otter.ai charges ~€20/month. Dragon NaturallySpeaking ~€500 one-time. Open Speech Studio is free forever - open source LGPL-3.0.

€0 forever
♾️

No rate limits

Dictate 10 hours a day if you want. No API quotas, no 'Fair Use Policy', no surprise invoices. Your hardware is the only limit.

Feature Open Speech Studio Dragon NaturallySpeaking Otter.ai
Price €0 (open source) ~€500 one-time ~€20/maand
100% local Yes Yes No (cloud)
GDPR-safe without DPA Yes Yes No
Languages 99 talen (Whisper) ~6 talen ~10 talen
GPU acceleration (CUDA) Yes No N/A (cloud)
Open source Ja (LGPL-3.0) No No

Features - what can it do?

Open Speech Studio is more than a dictation tool. It's a complete speech-to-text suite with audio file transcription, meeting recordings, multiple Whisper models, and export to standard subtitle formats.

🎙️

Dictate anywhere

Dictate text directly into any application via a system-wide overlay. Works in Outlook, Word, Chrome, VS Code, Slack, Teams - literally anywhere in your OS.

📂

Audio file transcription

Transcribe audio files (.mp3, .wav, .m4a, .ogg) with parallel processing. Multiple files at once, fast and accurate.

👥

Real-time meeting transcription

Record project meetings and watch the transcription appear in real time. Automatic minutes, even offline on the job site.

CUDA GPU acceleration

CUDA GPU acceleration for faster results. On a modern GPU, Whisper Large runs nearly real-time - optimal use of your hardware.

⌨️

Keyboard shortcuts

CTRL+Win for the overlay, CTRL+SHIFT+Space as alternative. No mouse click needed, no menu hunting - dictation starts in <100ms.

🧠

Whisper AI models

5 models: tiny, base, small, medium, large. Choose between speed (tiny) or accuracy (large). Memory usage 1-10GB depending on model.

🌍

99 languages

Whisper supports 99 languages: Dutch, English, German, French, Spanish, Polish, Arabic, Mandarin and many more. Automatic language detection.

🎛️

Spotify-style overlay

Sleek, minimalist overlay UI with live waveform visualization. Stays out of the way, disappears automatically after inserting.

🕘

History

Full history of transcriptions - searchable, reusable. Everything stored locally, you stay in control.

📄

Export TXT, SRT, VTT

Export transcriptions as plain text (.txt) or as subtitles (.srt, .vtt) for video content with timestamps.

📋

Clipboard integration

Automatic clipboard integration - dictate and paste directly, or let text auto-insert at the cursor position.

🔒

100% local

Everything runs locally on your own computer. Your data never leaves your PC. No cloud, no subscription, no limits. GDPR-safe out of the box.

Keyboard shortcuts

Start speech-to-text anywhere in your operating system:

CTRL + Windows or CTRL + SHIFT + Spatie

Configurable in settings - choose your own shortcut if it conflicts with other software.

Tech stack

Open Speech Studio is built on a modern, performant stack. OpenAI Whisper for state-of-the-art speech recognition. Rust for a lightweight, fast backend (no Python runtime needed). Tauri 2 for the desktop app - much lighter than Electron (~10MB instead of 200MB). TypeScript + Svelte for the UI. Fully open source under LGPL-3.0.

Download now - 30 second setup

1 installer file (.exe), no registration, no account, no credit card. Ready to dictate in <1 minute. Works on Windows 10 and Windows 11.

v0.10.3 stable · LGPL-3.0 · Windows 10/11 · ~40MB download