Changelog

Ogni release, miglioramento e correzione di bug — direttamente dalle nostre release di GitHub.

v1.8.313 agosto 2026

1.8.3

# OpenWhispr 1.8.3 > Note: 1.8.2 was never published — this release includes everything from it. If you're updating from 1.8.1, all of the below is new to you.

🌟 Highlights

🖊 Edit highlighted text by voice

Select text in any app, trigger the voice agent, and say what you want changed — "make this more formal", "fix the grammar", "turn this into bullet points". The agent rewrites the selection in place instead of pasting a new block. Works even in apps whose accessibility trees stay dormant (Arc, Chrome, Slack, VS Code, Claude Desktop, and more).

🖥 Screen context for the voice agent

Turn on "Share screen context" (Settings → AI Models → Voice Agent, off by default) and the agent sees the display you're working on when you speak — so "reply to this email" or "explain the error on screen" just works. You can route screenshot-carrying commands to a dedicated vision model, and a screenshot that can't be sent never costs you the command. Screenshots live in memory for one request only — never written to disk, stored, or logged.

📅 Microsoft & Apple Calendar

Outlook and Apple Calendar join Google Calendar: your meetings are detected, reminders fire before they start, and meeting notes link to the right event. Google Calendar sync also now handles large calendars completely.

✨ Also new (from 1.8.2)

  • Collaboration is free — Any signed-in account can create and join team spaces, share individual notes, and sync shared content. Paid workspaces see prorated seat costs before inviting.
  • Managed Enterprise AI — Organizations can centrally configure Amazon Bedrock or Azure OpenAI for cleanup, voice agent, notes, and translation. Employees sign in with company SSO — no cloud keys to handle — and prompts go directly to the organization's cloud account.
  • Organization policy, fully enforced — Policy now applies across the whole app: restricted options are hidden with safe fallbacks, and enforcement covers modes, providers, features, sharing, and retention.
  • Tinfoil realtime meeting transcription — Tinfoil joins the realtime meeting providers.
  • Agent failures are no longer silent — If the agent can't process a command, you get an "Agent Unavailable" notice instead of your raw words pasting into whatever app you were in.

⚡ GPU acceleration you can trust (new in 1.8.3)

  • "GPU acceleration active" now means it. The indicator reflects what the transcription engine is actually running on — ready, activating, active, or "could not be activated" with a Retry — instead of turning green whenever a download finished.
  • Enable GPU works instantly. Downloading or removing a GPU pack applies immediately — no app restart, no silent CPU fallback.
  • A GPU failure never costs you a dictation. If the GPU engine crashes, the same recording is transcribed on CPU and pasted, and the failed backend is remembered instead of being retried on every launch.
  • GTX 10-series (Pascal) cards can now use CUDA. The CUDA pack ships Pascal kernels for the first time; cards the build can't run on (Maxwell and older) are offered the Vulkan pack that works — no more loading the model and crashing at first use.
  • GPU packs can no longer corrupt each other. Each pack installs into its own directory, installs are atomic (a power cut can't leave a half-installed pack), and old installs are healed automatically.

🔒 Your keys and audio stay where you point them (new in 1.8.3)

  • Custom endpoints fail closed. A missing or invalid custom URL for speech-to-text or AI cleanup now fails with a clear error instead of silently sending your audio, prompt, and API key to OpenAI. ⚠️ Action required: if you use the Custom provider and never changed its pre-filled URL, set a real endpoint under Settings — or switch to the OpenAI provider.
  • Keys are scoped and stored securely. Your cleanup key no longer rides along to other endpoints, and custom-endpoint keys moved from plaintext into the OS secure store (migrated automatically).
  • Every provider remembers its model. Switching providers and back — for transcription and all six AI scopes — restores your previous choice instead of resetting it.
  • Transcription errors tell the truth. A broken transcription engine now reports a real error and keeps the recording for retry, instead of blaming your microphone with "No Audio Detected."

🛠 Fixes & polish

  • Dictation — No more clipped first words, realtime streaming waits for your final words instead of racing a timer, voice activity detection is now opt-in, dropped transcript segments are retried and surfaced, and transcripts are no longer replaced unexpectedly.
  • Meetings & speakers — Speaker labels you set stay put across live and offline identification, live speaker identification lands on Windows, your own dictation no longer triggers meeting detection, diarization writes land on the right note, meeting prompts are steadier, and realtime streams stop cleanly. Intel Mac meetings no longer fail at startup.
  • Notes & sync — Cloud sync can't overwrite your local edits with an empty copy, deleted notes can't resurrect, edits made mid-save stay pending, and idle collaboration sync backs off politely.
  • Launch at login on Linux — With correct behavior across GNOME and KDE, joining macOS and Windows (which also got start-hidden fixes).
  • macOS — The Globe hotkey no longer also opens the system emoji/input switcher.
  • Linux — Meeting notifications stay clickable, and the text monitor builds correctly against AT-SPI2.
  • Interface — Markdown exports get correct timestamps, empty states close their gaps, the "Coming Soon" badge translates again, and hotkeys with left/right modifiers parse correctly.

🙏 Thanks

Thanks to @Chadpiha, @hsusul, @xAlcahest, @greatcoat, @boseq, and @stantheman0128 for their work on this release — and a warm welcome to first-time contributors @iSparsh and @edwin-luu! 🎉
Visualizza su GitHub →
linux-text-monitor-v1.0.012 agosto 2026

Linux Text Monitor v1.0.0

Prebuilt Linux text monitor binary for auto-learn correction monitoring.
This binary uses AT-SPI2 to detect text field changes after paste.
Visualizza su GitHub →
windows-fast-paste-v2.0.06 agosto 2026

Windows Fast Paste v2.0.0

Prebuilt Windows fast-paste binary for clipboard paste and selection-copy operations.
Uses Win32 SendInput API with automatic terminal detection for paste (Ctrl+V/Ctrl+Shift+V) and copy (Ctrl+C/Ctrl+Shift+C).
Visualizza su GitHub →
v1.8.130 luglio 2026

1.8.1

# OpenWhispr 1.8.1 Everything from 1.8.0 — team spaces, web note sharing, new local models — plus a critical fix. If you installed 1.8.0, please update right away.

🚑 Fixed in 1.8.1

  • Note content could leak between notes. In 1.8.0, switching from one note to another could copy the second note's content into the first, even if you didn't edit either one. This is fixed — edits now stay with the note they belong to. (If you used 1.8.0, it's worth a quick look through recently-viewed notes.)

✨ What's new (from 1.8.0)

  • Team spaces — Create shared spaces alongside your private notes, invite teammates by email with roles, and jump straight in from a one-click link. Your notes stay on your device and sync in the background, with access checked on the server so only the right people can see them.
  • Share notes on the web — Publish a note to notes.openwhispr.com for anyone with the link, anyone in your company's email domain, or just the people you invite.
  • Workspace settings — Manage members, team spaces, and billing from the new Settings → Workspace section.
  • New local AI models — Added the Gemma 4 family of on-device models for faster, fully private AI.
  • Privacy retention controls — Choose how long saved recordings and transcripts are kept — down to a single day — and OpenWhispr cleans up the rest automatically.
  • See your raw transcript — History now shows the full, unedited transcript with a one-click copy button.
  • Dictionary from the command line — List and bulk-update your custom dictionary through the CLI.

🛠 Fixes & polish (from 1.8.0)

  • More accurate dictation — Local transcription no longer chops words in half, long recordings don't time out early, and a dropped microphone no longer loses what you just said.
  • Steadier cloud & streaming transcription — Sturdier large-file uploads, cleaner NVIDIA Parakeet and AssemblyAI streaming, and your transcript is kept even if the cleanup step comes back empty.
  • Safer notes sync — Cloud sync will no longer overwrite a note you've edited locally with an empty copy.
  • Better exports — Meeting participants are included, subtitle timing is fixed, and speaker-less segments are merged more cleanly.
  • Translation — Local translation is more reliable and handles empty results gracefully.
  • Windows — No more console windows flashing during dictation, meetings, and GPU checks.
  • Linux — More reliable pasting and meeting detection, and better behavior on Wayland.
  • macOS — The Dock icon now follows the control panel, the tray icon toggles it, and the Accessibility setup handles the app not being listed yet.

🙏 Thanks

Thanks to @Chadpiha, @hsusul, @xAlcahest, @IdrisGit, @greatcoat, @dpersek, and @sochotnicky for their work on this release — and a warm welcome to first-time contributors @stantheman0128, @dreasan, @fabrizio2210, @William-Ger, @ambbone, @AIalliAI, and @B3rK-3! 🎉
Visualizza su GitHub →
v1.8.013 agosto 2026

1.8.0

# OpenWhispr 1.8.0 Work together in shared team spaces and share notes on the web — plus new local AI models, privacy retention controls, and a big reliability pass across dictation and notes.

✨ What's new

  • Team spaces — Create shared spaces alongside your private notes, invite teammates by email with roles, and jump straight in from a one-click link. Your notes stay on your device and sync in the background, with access checked on the server so only the right people can see them.
  • Share notes on the web — Publish a note to notes.openwhispr.com for anyone with the link, anyone in your company's email domain, or just the people you invite.
  • Workspace settings — Manage members, team spaces, and billing from the new Settings → Workspace section.
  • New local AI models — Added the Gemma 4 family of on-device models for faster, fully private AI.
  • Privacy retention controls — Choose how long saved recordings and transcripts are kept — down to a single day — and OpenWhispr cleans up the rest automatically.
  • See your raw transcript — History now shows the full, unedited transcript with a one-click copy button.
  • Dictionary from the command line — List and bulk-update your custom dictionary through the CLI.

🛠 Fixes & polish

  • More accurate dictation — Local transcription no longer chops words in half, long recordings don't time out early, and a dropped microphone no longer loses what you just said.
  • Steadier cloud & streaming transcription — Sturdier large-file uploads, cleaner NVIDIA Parakeet and AssemblyAI streaming, and your transcript is kept even if the cleanup step comes back empty.
  • Safer notes sync — Cloud sync will no longer overwrite a note you've edited locally with an empty copy.
  • Better exports — Meeting participants are included, subtitle timing is fixed, and speaker-less segments are merged more cleanly.
  • Translation — Local translation is more reliable and handles empty results gracefully.
  • Windows — No more console windows flashing during dictation, meetings, and GPU checks.
  • Linux — More reliable pasting and meeting detection, and better behavior on Wayland.
  • macOS — The Dock icon now follows the control panel, the tray icon toggles it, and the Accessibility setup handles the app not being listed yet.

🙏 Thanks

Thanks to @Chadpiha, @hsusul, @xAlcahest, @IdrisGit, @greatcoat, @dpersek, and @sochotnicky for their work on this release — and a warm welcome to first-time contributors @stantheman0128, @dreasan, @fabrizio2210, @William-Ger, @ambbone, @AIalliAI, and @B3rK-3! 🎉
Visualizza su GitHub →
v1.7.618 luglio 2026

1.7.6

# OpenWhispr 1.7.6 Dictation translation, audio import with speaker detection, GPU support for AMD and Intel, live streaming transcription — and a deep round of self-hosted and reliability fixes.

New

  • Translation mode — dictate in any language and paste the text in another. Its own hotkey, up to 5 target languages, and a dedicated model, configured under Settings → AI Models → Translation.
  • Import audio from URLs and in batches — paste a YouTube or direct audio link, queue up to 50 URLs and files, and turn on Speaker detection to label who said what. Detection runs on-device by default.
  • Vulkan GPU acceleration — the one-click local Whisper GPU flow now covers AMD Radeon and Intel Arc/integrated GPUs on Windows and Linux, with automatic CPU fallback.
  • NVIDIA Nemotron streaming models — live text as you speak from a persistent local stream, and dictation now commits the streamed text the moment you stop (no second decode, roughly half the CPU per dictation).
  • Liquid AI LFM2/LFM2.5 — five new local reasoning models for on-device cleanup, down to a 0.25 GB model that runs on modest hardware.
  • Collapsible sidebar — collapse it for full-width notes; hover the left edge to peek.
  • Meeting prompts, unified — calendar reminders now use the in-app overlay (they survive Focus/Do Not Disturb and never appear in screen shares), with one-click Join & transcribe when the event has a meeting link.

Fixed

  • Self-hosted: audio uploads, History retries, note formatting, and Chat all respect your self-hosted server now — no more silent fallback to a cloud provider — plus model lists work without /v1 and "disable thinking" works on Ollama.
  • Local AI: starting a local cleanup model no longer freezes the whole machine — the llama server runs with a bounded context, and a model that truly doesn't fit fails with a visible error.
  • Cleanup providers: Mistral works as a custom provider instead of failing with a 422, Groq no longer fails silently, and if cleanup ever fails your dictation is pasted raw with a toast instead of being lost.
  • macOS: pausing your music during dictation works again on macOS 15.4+, and Fn+Arrow shortcuts no longer trigger stray push-to-talk recordings.
  • Microphones: a muted or vanished mic falls back to the default device instead of recording silence, and your selection survives Chromium's device-ID rotation.
  • Windows: Parakeet installs no longer hang behind an old PATH tar, and the event-driven mic listener that powers meeting detection now actually ships.
  • Linux: .deb installs and upgrades no longer fail in containers — and upgrading no longer deletes your downloaded models.
  • Notes: search works in every language (Cyrillic, CJK, Arabic, accented text), and Generate Notes regenerates auto-assigned titles while leaving yours alone.
Plus many more — full list in the changelog.

Thanks

Thanks to @xAlcahest, @Chadpiha, and @sgrimbly for their work on this release — and a warm welcome to first-time contributors @hsusul, @sochotnicky, @cgkades, and @Serhii-Leniv!
Visualizza su GitHub →
windows-mic-listener-v1.0.014 luglio 2026

Windows Mic Listener v1.0.0

Prebuilt Windows mic listener binary for event-driven microphone detection.
Uses WASAPI IAudioSessionManager2 to monitor capture sessions. Supports --exclude-pid to filter out OpenWhispr's own mic usage.
Visualizza su GitHub →
v1.7.511 luglio 2026

1.7.5

# OpenWhispr 1.7.5 New models and providers, multiple hotkeys per action, and around a dozen fixes across meetings, notes, transcription, and every platform.

New

  • Latest cloud models: OpenAI GPT-5.6 (Sol/Terra/Luna) and Anthropic Claude Fable 5 and Sonnet 5 — plus Claude Fable 5 for AWS Bedrock.
  • OpenRouter is now a built-in provider, with a searchable, grouped picker for its 300+ models.
  • Corti can run AI cleanup and reasoning, not just transcription. For healthcare use, that text stays on Corti (EU) or the HIPAA-compliant OpenWhispr Cloud — never a third-party model.
  • Tinfoil confidential transcription now covers uploaded audio, not just live dictation.
  • Multiple hotkeys per action — bind more than one key to dictation, agent, voice agent, or meeting.
  • Enterprise / AWS Bedrock: region-aware model IDs, a live model catalog, and Agent Mode (tool-calling) support.

Fixed

  • Meetings: remote speakers no longer collapse into one; long calls reconnect past the 60-minute limit without losing audio; fewer words dropped during cross-talk.
  • Notes: enhancement no longer reuses another note's transcript or renames notes you've already titled.
  • Cloud transcription: long recordings retry failed chunks instead of failing or leaving gaps.
  • macOS: dictation into Claude Desktop and claude.ai no longer breaks after the first paste.
  • Windows: fixed a false "No audio detected", the firewall prompt for local Parakeet, and local transcription failing on some builds.
  • Linux: fixed Wayland global-hotkey crashes on Node 24.
Plus several smaller fixes — full list in the changelog.

Thanks

Thanks to @xAlcahest, @danielmccannsayles, @IdrisGit, and @Alexey-Sachko for their work on this release — and a welcome to @Alexey-Sachko on their first contribution!
Visualizza su GitHub →
v1.7.47 luglio 2026

1.7.4

# OpenWhispr 1.7.4 A feature-and-reliability release: two new privacy-first ways to transcribe and reason, native Windows meeting audio, sync that keeps working in the background, and a broad sweep of fixes across dictation, clipboard, updates, and Linux — plus security hardening. Nothing here changes your existing setup; it all just works better.

Highlights

  • Tinfoil — confidential cloud (BYOK). A new provider that runs both transcription *and* AI cleanup/agent inside attested secure enclaves, so your audio and text stay private even when processed in the cloud. (#875, #944)
  • Azure AI Foundry / Azure OpenAI speech-to-text. Point the custom transcription provider at your own Azure deployment — OpenWhispr now speaks Azure's endpoint format and auth. (#997)
  • Native Windows meeting audio. Meetings now capture sound from every app on every output device instead of only your default speakers, via a native Windows helper. Older Windows automatically falls back to the previous method. (#960)
  • Enterprise SSO at sign-in, right from onboarding. (#1034)

Sync

  • Runs in the background, even from the tray — dictations upload and other devices' changes pull down on window focus, on network reconnect, and every 5 minutes, even if you never open the main window. (#1070, #1072)
  • Snippets now sync across your devices, like your dictionary already does. (#1037)
  • No more duplicate folders when you sign in on a new device. (#1086)

Dictation & notes

  • Cleaner cleanup — your transcript gets tidied instead of occasionally getting "answered," and dictations only go to your voice agent when you actually address it by name. (#1073)
  • Turkish snippets — triggers containing İ / ı now match correctly. (#1050)
  • API keys no longer vanish when you type one and click away — a papercut that mostly hit Corti/BYOK onboarding. (#1039)

Clipboard, audio & media

  • Rich clipboard formats survive auto-paste, clipboard restore is faster, and pasting reliably targets the right app on macOS. (#1020, #1038, #1000)
  • Mic falls back to your default device when a saved one goes stale, and empty recordings no longer crash transcription. (#978, #891)
  • Paused music/video resumes the instant you stop recording (not after transcription) and won't get stuck paused on a quick tap. (#1030, #1061)

Platform & reliability

  • Updates: fixed a crash on the first "Install & Restart" click after an update — macOS installs no longer stall. (#1012)
  • Local transcription: GPU model reloads after sleep, multi-GPU machines pick the right card, the voice-activity model now ships in packaged builds, realtime streaming is hardened, and llama.cpp was updated. (#1032, #1018, #1000, #1044, #995)
  • Hotkeys: modifier-only hotkeys (dictation, voice agent, agent, meeting) now work simultaneously on Windows/Linux. (#1001)
  • Linux: installs on openSUSE and launches on hardened kernels that restrict user namespaces. (#1014, #1042)
  • Self-hosted: transcription endpoints can now choose their own model. (#1043)
  • Quieter startup — suppressed benign Qdrant warnings. (#1083)

Security

  • Cleared all critical/high dependency alerts and restored lockfile integrity with a CI guard (SOC 2 secure-code hardening). (#1051, #1069)

Contributors

Thanks to everyone who shipped 1.7.4 — with a warm welcome to our first-time contributors (*): @xAlcahest · @Chadpiha · @IdrisGit · @danielmccannsayles* · @zachdotai* · @bikramjitk* · @SyntaxSawdust* · @UmutEmreOnder* · @greatcoat* · @chirag127*
Visualizza su GitHub →
windows-system-audio-helper-v1.0.07 luglio 2026

Windows System Audio Helper v1.0.0

Prebuilt Windows system audio helper binary for meeting transcription.
Captures system audio via WASAPI process loopback (exclude mode), hearing every application on every output device while excluding OpenWhispr's own audio. Requires Windows 10 2004 or later; older systems automatically fall back to Chromium display-media loopback.
Visualizza su GitHub →
v1.7.324 giugno 2026

1.7.3

✨ New
  • Snippets — say a shortcut like "cal link" and the full text gets pasted
  • Voice Agent hotkey — one key sends your speech straight to your AI assistant
  • Medical dictation via Corti for clinical-grade accuracy
  • Smarter setup that adapts to how *you'll* actually use OpenWhispr
  • Redesigned dictionary that now syncs across your devices
  • Audio upload gets its own settings and a cancel button
  • Discarded dictations are saved, so nothing gets lost
  • Notification controls to pick which alerts OpenWhispr can show
  • New AI models: Claude Opus 4.8, Gemini 3.5 Flash, Gemma 4, plus xAI transcription
Visualizza su GitHub →
v1.7.220 maggio 2026

1.7.2

# OpenWhispr 1.7.2 A small patch on top of 1.7.1 — zero unnecessary Keychain prompts on first launch, cloud transcription working again on Electron's net.fetch, and the Note Formatting selector now actually controls model routing.

Features

  • Note Formatting selector now routes Generate Notes. The Note Formatting tab's model / provider / mode selectors were a no-op end-to-end in 1.7.1 — every Generate Notes call silently fell through to the Cleanup model. Now wired through the same per-scope plumbing, with a Cleanup fallback so untouched-settings users keep their current behavior.
  • Wave Terminal paste. Wave Terminal is now in the Linux terminal allowlist, so auto-paste routes through Ctrl+Shift+V instead of the default Ctrl+V.

Bug Fixes

  • No more spurious macOS Keychain prompts on first launch. First-launch Keychain prompts drop from ~3 to 0. The secret-crypto backend no longer eagerly probes itself before any window appears — it defers Keychain access until you actually save your first secret. The safeStorage key backup is also only written when a new master key is generated, not on every launch.
  • Cloud transcription works again. The 1.7.0 migration from https.request to Electron's net.fetch carried over a manual Content-Length header, which Electron rejects as a forbidden Fetch header — failing with net::ERR_INVALID_ARGUMENT before any bytes hit the wire. Fix covers all five upload paths: cloud transcribe, chunked cloud transcribe, retry, file upload, and BYOK whisper-compatible.
  • Notes view state stability. Fixed a stale-ref issue where switching between notes could lose unsaved enhanced-content edits.
Visualizza su GitHub →
v1.7.120 maggio 2026

1.7.1

Features

  • macOS mouse-button hotkeys for dictation
  • Google Calendar: new "sync primary calendar only" toggle to skip shared ones
  • Make automatic note titling optional in Settings
  • Smarter local transcription that ignores silence and breathing, fully configurable for local models

Bug Fixes

  • Cloud meeting transcription restored after a recent OpenAI Realtime API change
  • Local Whisper transcription restored on Windows
  • Music auto-pause and resume works again on the latest macOS
  • Self-hosted models now use their own URLs and API keys instead of falling back to cleanup settings
  • Custom dictionary now reaches meeting note cleanup
  • Language preference preserved on retry, in meetings, and in the dictation preview
  • Custom prompts sync across windows instantly without restart
  • Transcription sync no longer crashes on bad cloud data
  • Idle local AI models release GPU memory after a timeout
  • Works correctly behind corporate proxies for model downloads, GPU detection, and calendar sync
  • Linux: Pop!OS COSMIC support, GNOME Wayland terminal auto-paste, Konsole paste fix, and the dictation overlay no longer steals focus on Sway/i3/Hyprland
Visualizza su GitHub →
v1.7.04 maggio 2026

1.7.0

Breaking for Mac users: please delete your old OpenWhispr first. 1.7.0 uses a new app ID — we had to migrate Apple Developer accounts, so you'll be asked to re-grant permissions on first launch. Notes, settings, API keys, and downloaded models carry over automatically.

Summary

  • More sign-in options — new: Sign in with Microsoft, Sign in with Apple on macOS. Auth runs on our own infra now (migrated to Better Auth); self-hosters can point at their own server.
  • Sessions in your OS keychain — survive crashes and Electron restarts. No more random sign-outs.
  • API keys encrypted at rest — all 12 BYOK + enterprise creds moved from .env to the OS keychain. Silent one-time migration on first launch.
  • Background meeting recording — navigate to other notes, open Settings, switch tabs; the audio pipeline keeps going. New floating pill shows live mic levels and clicks back to the recording note.
  • Per-note diarization preferences persist across stop/resume.
  • Cleaner mic capture with a new acoustic gate + better echo cancellation. Music pause/resume on Windows works again.
  • Per-scope LLM setup — pick different providers/models for cleanup, agent, formatting, and chat. You can now have text cleanup toggled ON but your dictation agent toggled OFF.
  • NVIDIA Parakeet parakeet-unified-en-0.6b — new English-only model, 5.91% avg WER, ~631 MB.
  • openwhispr CLI talks directly to the desktop app via a local HTTP bridge (127.0.0.1, bearer-token); falls back to cloud when the app is closed.
  • Proxy-aware fetches + OS CA trust — corporate TLS interception works.
  • Auto-learn corrections now work for Cyrillic, CJK, Arabic, Devanagari.
  • A bunch of other small bug fixes and performance improvements.

Detailed changelog

  • build(mac): migrate Apple signing to Gizmo Labs Inc. by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/639
  • fix(qdrant): set cwd to STORAGE_DIR so packaged app doesn't crash on startup by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/640
  • feat(sync): cross-device delete propagation by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/646
  • feat(meeting): interop cloud streaming providers + AEC/VAD improvements by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/656
  • Re-grant permissions modal for 1.6.11 upgraders by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/659
  • fix(chat): persist first user message when creating a new conversation by @xAlcahest in https://github.com/OpenWhispr/openwhispr/pull/662
  • fix(tls): trust OS CA store for Node-side TLS; keep realtime WS on OpenAI only by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/657
  • Fix adding non-ASCII languages to dictionary by @wake0up0ne0 in https://github.com/OpenWhispr/openwhispr/pull/666
  • Fix dictionary correction tracking for TextPattern-based controls (alongside existing ValuePattern support) by @wake0up0ne0 in https://github.com/OpenWhispr/openwhispr/pull/665
  • Bundle ID migration: com.herotools.openwispr → com.gizmolabs.openwhispr by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/669
  • feat(cli): local HTTP bridge for unified CLI by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/676
  • fix(notes): update folder selection when moving notes between folders by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/678
  • refactor: split dictation cleanup from dictation agent + per-scope LLM config by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/677
  • refactor(auth): switch desktop to Better Auth + add Microsoft sign-in by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/686
  • fix(network): proxy-aware fetches + actionable connectivity errors by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/687
  • chore(release): 1.7.0 confidence cleanup by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/689
  • ci(release): rename VITE_NEON_AUTH_URL → VITE_AUTH_URL by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/690
  • fix(auth): initiate desktop OAuth from browser, not renderer by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/692
  • feat(auth): Sign in with Apple on macOS by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/691
  • fix(onnx): cap embedding segments and isolate inference in utility process by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/693
  • fix(sidecars): reap children on quit and on next launch (#683) by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/694
  • docs(changelog): plain-English 1.7.0 + lockfile refresh by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/695
  • fix(llama): raise Vulkan startup timeout, stop server before re-download by @xAlcahest in https://github.com/OpenWhispr/openwhispr/pull/698
  • fix(media): fall back to media key when GSMTC fails on Windows by @xAlcahest in https://github.com/OpenWhispr/openwhispr/pull/697
  • docs(readme): acknowledge Hugging Face as model hub by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/703
  • feat(transcribe): generate clientTranscriptionId for cloud sync dedup by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/702
  • feat(security): encrypt API keys at rest via safeStorage by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/629
  • fix(media): use AsTask bridge for WinRT async calls in GSMTC scripts by @xAlcahest in https://github.com/OpenWhispr/openwhispr/pull/706
  • fix(media): resume playback on no-audio-detected event by @kdenney in https://github.com/OpenWhispr/openwhispr/pull/701
  • feat(auth): switch desktop to Authorization: Bearer + safeStorage by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/704
  • fix(ci): compile macos-media-remote binary in release and build workflows by @kdenney in https://github.com/OpenWhispr/openwhispr/pull/700
  • feat(reasoning): self-hosted OpenAI-compatible parity + thinking-mode toggle by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/708
  • feat(parakeet): add parakeet-unified-en-0.6b model by @milanleonard in https://github.com/OpenWhispr/openwhispr/pull/713
  • feat(meeting): background recording with floating pill and responsive layout by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/709
  • fix(runtime): unblock model downloads, surface ONNX worker errors, dedupe hotkey re-register by @gabrielste1n in https://github.com/OpenWhispr/openwhispr/pull/716
  • fix(linux): resolve symlink in launcher wrapper by @xAlcahest in https://github.com/OpenWhispr/openwhispr/pull/717

New Contributors

  • @wake0up0ne0 made their first contribution in https://github.com/OpenWhispr/openwhispr/pull/666
  • @kdenney made their first contribution in https://github.com/OpenWhispr/openwhispr/pull/701
  • @milanleonard made their first contribution in https://github.com/OpenWhispr/openwhispr/pull/713
Visualizza su GitHub →
v1.6.1020 aprile 2026

1.6.10

Features

  • Speaker diarization controls — per-meeting toggle, "N others in call" stepper, auto-label on 1-on-1s, fallback to "You/Others" when labeling is off
  • Integrations hub — new top-level view with MCP connectivity for Claude, ChatGPT, and Cursor; API keys moved here
  • Enterprise connections - connect to your organisations instance via bedrock

Bug Fixes & Improvements

  • Settings reorganization — "AI Models" split into Speech-to-Text and Language Models with clearer sub-tabs
  • Agent hotkey moved under Hotkeys
  • Meeting transcription tuned to reduce GPT-4o hallucinations during silence
  • Model cache opens at the correct folder so downloaded models are actually visible
  • Sync — cleared transcripts and deleted folders no longer reappear after restart
  • Echo leak detector now runs pre-AEC so it can actually see system audio bleed
Visualizza su GitHub →
v1.6.916 aprile 2026

1.6.9

New features

Transcript Export — Save transcripts to disk as TXT, SRT, or JSON directly from the meeting view. Cloud Sync — Notes, folders, conversations, and transcriptions now sync bidirectionally across devices. API Keys — Create and manage API keys from Settings to integrate OpenWhispr with your workflows programmatically. Linux Push-to-Talk — Native push-to-talk via evdev with guided permission setup. Auto-Paste Toggle — New setting to disable automatic pasting after dictation. Agent Folders — The agent can now list, create, and match folders semantically when organizing notes.

Improvements

  • Upgraded to Electron 41 and Node 24
  • Redesigned provider tabs as compact pill buttons
  • Local model descriptions replaced with clickable spec links
  • llama.cpp detection now probes /v1/models and prefers /chat/completions

Fixes

  • Windows hotkey no longer false-triggers after Win+L lock/unlock
  • Fixed stuck recording when push-to-talk key-up is missed on Windows
  • Relaxed speech gate thresholds with no-audio toast for failed transcriptions
  • Retry button now shown on failed transcriptions
  • Notes sync when updated externally
  • Tuned speaker diarization to prevent excessive speaker creation
  • TLS/certificate errors now surfaced in model download UI
  • Floating icon position persists across restarts
  • Fixed notification window sizing and dismiss circle visibility
Visualizza su GitHub →
v1.6.815 aprile 2026

1.6.8

New features

  • Meeting recording with on-device speaker diarization — capture mic and system audio together, and every segment is automatically labeled with who said it. Diarization runs 100% locally: voices never leave your device. System audio capture now works on Windows and Linux (Chromium loopback and the native audio portal), not just macOS.
  • Local live transcription preview — see your words appear in real time as you speak, streamed from a local Whisper or Parakeet model before you release the hotkey. Fully on-device, no cloud round trip.
  • Self-hosted support — first-class option to point OpenWhispr at your own LAN server for both transcription and agent inference. One click to swap between OpenAI, Anthropic, Gemini, a local GGUF, or your self-hosted endpoint.
  • Speaker to contact linking — voice profiles bind to calendar attendees, so recurring participants get named automatically after the first tag.
  • Improved echo cancellation — WebRTC AEC now runs as a native sidecar, paired with a custom cross-correlation detector and text-level dedupe to cleanly handle external speakers, system-audio capture, and low-echo-quality mics.
  • Gemma 4 models — Gemma 4 31B and Gemma 4 26B MoE added to the local model registry.
  • Smart transcription error handling — when a transcription fails, inline CTAs let you retry, switch provider, or open logs.

Linux

  • Hyprland global hotkeys via hyprctl, plus correct active-window detection for terminal paste.
  • GNOME shortcut conflict detection, active-key display, and proper fallback when the requested shortcut is taken.

Fixes

  • Windows: installer DPI scaling corrected, GPU compositing re-enabled.
  • Local Whisper: stricter speech gate cuts hallucinated output on silent audio.
  • Diarization: session-scoped results, disk spooling for long meetings, reliable spinner, correct speaker label propagation.
  • Notes: transcripts scope to their originating note, view auto-switches when a meeting recording starts.
  • UI: fixed 420px window width that grows only with text, consistent across phases; preview overlay redesigned.
  • Settings: inference mode auto-switches when you select a new provider.
  • i18n: speaker labels, cloud settings, integrations, and calendar strings tightened across locales.
  • Stability: 22 React Compiler lint issues resolved, tighter effect dependencies, cleaner hook usage across 15 files.
Visualizza su GitHub →
meeting-aec-helper-v1.0.014 aprile 2026

Meeting AEC Helper v1.0.0

Prebuilt WebRTC Audio Processing sidecar for meeting mic echo cancellation.
Binaries built from native/meeting-aec-helper/ against pinned webrtc-apm + abseil-cpp commits.
Visualizza su GitHub →
v1.6.73 aprile 2026

1.6.7

AI Chat & Semantic Search — Ask questions about your notes with the new embedded chat panel. Conversations sync to the cloud, and a local semantic search engine finds notes by meaning, not just keywords. Meeting Improvements — Calendar attendees automatically appear on meeting notes, auto-detection works for browser meetings, and echo cancellation cleans up mic input. Save Notes as Files — Export your notes to local Markdown files that mirror your folder structure. Responsive Settings — Settings dialog adapts gracefully to smaller windows. Bug Fixes — Resolved issues with meeting transcription, folder switching, clipboard pasting, and participant saving. Improved Linux Wayland support and Windows build signing. Huge thanks to @xAlcahest @DamianPala and others for their contributions to platform stability and improvements!
Visualizza su GitHub →
v1.6.623 marzo 2026

1.6.6

Rich Text Notes

  • Notes now use a rich text editor with Obsidian-style live preview — Markdown syntax hides as you type, giving you a clean writing experience

Meeting Transcription Upgrades

  • Dual-channel transcription — mic and system audio are captured separately with speaker-labeled chat bubbles so you can see who said what
  • Timestamped segments — meeting transcripts now include timestamps in chronological order
  • Smarter meeting notes — AI-generated summaries are now speaker-aware for more accurate meeting recaps

New Local Models

  • Mistral Nemo 12B and Gemma 3 12B are now available for on-device AI processing

macOS Audio Improvements

  • Native system audio capture via CoreAudio Tap — no more "screen recording" permission prompt on macOS 14.2+
  • macOS 15+ now shows the correct system audio consent dialog instead of the legacy screen recording one

Linux

  • KDE Wayland support — native global shortcuts now work on KDE Plasma via D-Bus, joining GNOME and Hyprland support
  • Fixed KDE Plasma overlay window and hotkey behavior
  • Fixed clipboard paste reliability on KDE Wayland

Simplified Permissions

  • Permission prompts consolidated to a single "Grant Access" button
  • Permissions are now re-checked against the OS each time you open settings, so the UI always reflects your actual state

Bug Fixes

  • Fixed Gemini agent streaming routing to the wrong endpoint
  • Fixed Windows mic volume being permanently altered during dictation
  • Fixed mono transcription failure on Linux
  • Fixed Bluetooth audio issues during meetings
  • Fixed meeting detection notifications firing when you're already in a meeting
  • Fixed held modifier keys not releasing before paste on Windows
  • Fixed paused media being unpaused during dictation
  • Fixed Google OAuth users skipping onboarding
  • Fixed AI cleanup prompt refusing to transcribe command-like speech
  • Fixed agent hotkey conflicts not showing a warning
Visualizza su GitHub →
v1.6.416 marzo 2026

1.6.4

Meeting Mode

  • Meeting mode hotkey — Assign a dedicated hotkey to instantly snap the panel into meeting mode and start a new meeting note
  • Smarter meeting detection — Reduced false positives; background apps like FaceTime no longer trigger notifications unless mic activity is detected
  • Meeting detection toggle — Disable meeting detection entirely from Settings

Auto-Update

  • Update notification — A slide-in notification appears when a new version is available, with an "Update Now" button

Multi-Monitor

  • Cursor-aware positioning — The floating icon now appears on whichever monitor your cursor is on

New Models

  • GPT-5.4 — Added as the new flagship OpenAI model
  • Qwen 3.5 — New local models added; removed sub-1B models

Bug Fixes

  • Fixed paused media resuming unexpectedly on Windows when no active playback sessions exist
  • Fixed Windows hotkey listener state corruption when switching hotkeys
  • Fixed silence detection rejecting valid speech (lowered threshold)
  • Fixed Windows paste not working in Windows Terminal (scan codes now included)
  • Fixed API keys saved in Settings being overridden by shell environment variables
  • Fixed local LLM models being deleted during Windows app updates
  • Fixed hotkey tooltip display for multi-key combos on macOS
  • Fixed RPM install conflicts with other Electron apps on Linux
  • Added Tailscale VPN (CGNAT range) support for self-hosted setups

Other

  • Panel start position now persists across app restarts
  • Cross-window settings sync (hotkey changes apply instantly)
  • Account deletion flow with cloud data cleanup
  • Agent mode window renamed to "Agent Chat"
Visualizza su GitHub →
v1.6.314 marzo 2026

1.6.3

Clearer Permissions

  • "Screen Recording" → "System Audio" — all permission prompts now clearly state we capture other participants' audio, not your screen
  • Electron 39 — eliminates the purple "screen recording" indicator, the "Your screen is being observed" lock screen message, and the misleading permission prompt on macOS 14.2+

Better Soft Voice Recognition

  • Auto Gain Control now enabled for dictation, automatically boosting quiet speech
  • Lower VAD sensitivity so soft-spoken audio is no longer missed
  • Less clipping at the start of speech — increased padding so quiet beginnings aren't cut off

Linux

  • Hyprland Wayland support — native global shortcuts using hyprctl keybindings + D-Bus

Bug Fixes

  • Fixed wl-copy failing silently on Wayland due to a too-short timeout
  • Fixed media staying paused after recording silence with "Pause media on dictation" enabled
Visualizza su GitHub →
v1.6.211 marzo 2026

1.6.2

System Audio Capture

Record system audio alongside your microphone in Notes — capture meeting audio, lectures, and more. Automatically enabled when screen recording permission is granted (macOS).

Smarter Meeting Detection

Meeting detection now uses native OS events instead of polling, cutting background CPU usage to near-zero.

Bug Fixes

  • Fixed hotkey reliability on Windows 11 (modifier-only shortcuts)
  • Fixed macOS Globe key failing silently on fresh installs — now prompts for accessibility permission
  • Fixed initial audio being dropped during realtime streaming
  • Fixed custom dictionary errors on large dictionaries
  • Fixed Parakeet model extraction on Windows 10
Visualizza su GitHub →
v1.6.19 marzo 2026

1.6.1

  • Even faster transcription on OpenWhispr Pro + Websockets for gpt4o
  • Bug fixes
Visualizza su GitHub →
v1.6.07 marzo 2026

1.6.0

  • Agent mode! Separate hotkey to spin up a chat, now you don't have to open a claude/chatgpt window. Check it out in settings.
  • Link your Google calendar and record meetings without a bot
  • Search your transcripts and notes
  • Brought back the ability to clear all transcripts
  • Choose voice recorder start position
  • Toggle to minimize the dashboard on startup
  • Toggle to pause music etc. when dictating
  • A bunch of other bug fixes and performance improvements
Massive thank you to @xAlcahest for all the major contributions to linux stability and other improvements!
Visualizza su GitHub →
windows-text-monitor-v1.0.026 febbraio 2026

Windows Text Monitor v1.0.0

Prebuilt Windows text monitor binary for auto-learn correction monitoring.
This binary uses Windows UI Automation to detect text field changes after paste.
Visualizza su GitHub →