Changelog
What's new in Whisperstream.
Every release, grouped and dated. The running record of what shipped and what changed.
- v1.1
Additional dictation languages
- Added support for 14 more dictation languages: Arabic, Cantonese, Chinese, Filipino, Hindi, Indonesian, Japanese, Korean, Macedonian, Malay, Persian, Thai, Turkish, and Vietnamese.
- AI cleanup and custom prompts now work with any OpenAI-compatible provider.
- Choose where speech models are downloaded.
- Start on login now works reliably.
- v1.0.1
Type into remote desktops and screen readers
- Keystroke input mode types your text out character by character instead of pasting it, so dictation works in remote sessions like Citrix and RDP and alongside screen readers.
- A new shortcut pastes your most recent transcription before AI cleanup, so the original is one keypress away.
- v1.0
Per-app cleanup and a private transcript history
- Automatic profile switching gives each app or website its own AI cleanup style. Whisperstream detects the active window and applies the right one.
- Transcript history keeps a searchable, private record of everything you dictate, with audio playback. It stays encrypted on your device.
- File import lets you drop in an existing audio file to transcribe it.
- Set separate shortcuts for push-to-talk and toggle recording.
- The free trial is now unlimited dictation for 7 days.
- v0.9
AI cleanup, now fully on-device
- AI Enhancement now runs on a built-in local model, on PCs with a capable graphics card. No API key, no cloud round trip.
- v0.8
AI cleanup, bring your own key
- Optional AI Enhancement sends transcriptions to Gemini, OpenAI, Anthropic, or a local Ollama model for cleanup before they're typed. It is off by default, and you bring your own API key.
- Pick Chat or Code mode, or write a custom style prompt.
- v0.7
Automatic updates and flexible shortcuts
- Automatic daily update checks, plus a Check for updates button on the About page.
- Auto-update now closes the app cleanly before installing.
- Your push-to-talk shortcut now accepts extra mouse buttons and modifier keys.
- Clearer errors when recording can't start, plus stability fixes.
- v0.6
Faster dictation
- Parakeet speech model now stays loaded between recordings, so dictation starts instantly.
- Faster app startup and close.
- v0.5
Signed installer and in-app Pro
- Whisperstream now ships with a verified Microsoft publisher signature, so Windows no longer warns about an unknown publisher.
- Buy and activate Pro without leaving the app.
- Click-through end user license agreement on first launch.
- Release notes now appear in the in-app update popup.
- v0.1
First release
- Push-to-talk dictation powered by NVIDIA Parakeet, running locally on your CPU.
- Live HUD waveform while you record.
- Custom dictionary for word and capitalization overrides.
- Usage dashboard with transcription, word, and recording-time stats.
- Auto-update via signed installer, with one-click log export for bug reports.