Is Wispr Flow Safe? What "Zero Data Retention" Actually Means
Wispr Flow processes your voice in the cloud, through named subprocessors — and its strongest privacy promise is a setting you have to find, not a default you get. Here's what to check, and four real local alternatives.
Short answer: Wispr Flow is a cloud dictation app. By its own Data Controls documentation, every recording you make is sent to third-party servers for processing — Baseten and Soniox handle the speech-to-text pipeline, OpenAI, Anthropic, and Cerebras process the resulting text, and everything runs through AWS infrastructure in the US. Wispr Flow states it plainly: Transcription always occurs on the cloud.
"Zero data retention" exists as a feature, which Wispr Flow calls Privacy Mode — but per its own subprocessors documentation, it's off by default for standard accounts, and there's currently no in-app way to opt out of the analytics and error-tracking that keep running regardless of Privacy Mode. None of this is hidden or unusual — Wispr Flow is more transparent about its architecture than most cloud dictation peers. But a documented cloud pipeline and a privacy policy are not the same thing as audio never leaving your machine. The question worth asking isn't really about one company. It's what to check before trusting any dictation app with your voice — and what your options are if the answer needs to be "nothing leaves my device."
What Does Wispr Flow Actually Do With Your Voice?
Per Wispr Flow's own Subprocessors & Third-Party Security page, here's the documented path your audio takes:
- Baseten and Soniox — run the transcription pipeline (speech recognition and formatting). Your raw audio is processed here.
- OpenAI, Anthropic, Cerebras — process the transcribed text for polish and formatting.
- AWS S3, us-east-1 — where dictation data is stored when Cloud Sync or standard mode is active.
- PostHog, Sentry, Segment — analytics and error tracking, which the same documentation confirms "there are currently no in-app controls to opt out" of.
The same page states plainly: Customers do not have individual approval rights over specific LLM providers or model families.
That's a reasonable trade-off for a cloud SaaS product built for speed and cross-device sync — the point isn't that this is unusual, it's that it is, definitionally, a cloud architecture.
Is "Zero Data Retention" a Default, or a Setting You Have to Find?
Wispr Flow's Privacy Mode is the mechanism behind "zero data retention": turn it on, and audio, transcripts, and edits stop being stored or used for training. The mechanics matter, though. Per Wispr's own Data Controls page: If you choose to disable "Privacy Mode," your Dictation Data may be used to evaluate, train and improve Flow's features and AI models.
Privacy Mode is off by default for individual users — it only locks on permanently if you sign the in-app HIPAA Business Associate Agreement, which is irreversible.
Worth knowing: even with Privacy Mode on or a BAA signed, Wispr's own subprocessors documentation is explicit that analytics and error-tracking tools keep running — signing the BAA "does not disable analytics or error tracking," and there's no toggle to turn that off separately.
Wispr Flow also ships an opt-in "Context Awareness" feature that reads on-screen content from your active app to improve formatting accuracy. It's disabled by default today. But it's the same feature that drew scrutiny in late 2025, when a Reddit thread titled "Wispr Flow Stores User Screenshots and is NOT Trustworthy" raised the concern publicly, followed by a related discussion in r/privacy. Whatever you conclude about that history, it's a useful reminder that a feature's default can change — which is exactly why checking the current settings yourself, on any app, beats trusting a headline.
What Should You Ask Any Dictation App About Privacy?
This isn't a Wispr Flow-specific checklist — it applies to every dictation tool, Inkvox included:
- Where does the audio actually go? — Does transcription happen on-device, or on a server somewhere? A vendor's own architecture docs (not its marketing page) will tell you.
- Who are the named subprocessors? — A legitimate cloud vendor publishes this list, as Wispr Flow does. If a company won't name who touches your audio and text, that silence is itself information.
- Is the protective setting the default, or something you have to enable? — "Off by default, available if you dig into Settings" is a very different posture from "on out of the box."
- Is it a policy, or an architecture? — A privacy policy can change with 30 days' notice. Audio that never leaves your device in the first place can't be retained by a future policy revision, because there was never a server in the path to retain it.
Real Local Alternatives to Wispr Flow
If your answer to "where does the audio go" needs to be "nowhere but my machine," here are real local-first tools people use — with their actual trade-offs, not just their pitch:
- OpenWhispr (macOS, Windows, Linux — open source, MIT) — runs Whisper or NVIDIA Parakeet locally via whisper.cpp with no telemetry, free for local use; an optional cloud/bring-your-own-key mode exists if you want it. Trade-off: it's a young project (2025), with a smaller install base than some peers, and output quality depends on which local model size you choose.
- Handy (macOS, Windows, Linux — open source, MIT, free) — built specifically to be, in its own words, "the most forkable" local dictation tool; every recording is processed on-device with no cloud mode at all. Trade-off: it's under active development and says so in its own README, which currently lists Whisper model crashes as a known "major issue, help wanted" and incomplete Wayland support on Linux.
- VoiceInk (macOS only — open source, GPL v3, $39.99 one-time) — local transcription via whisper.cpp, auditable source code, no subscription. Trade-off: no Windows or Linux build, and its AI text-cleanup features require bringing your own API key to a cloud LLM.
- Spokenly (macOS, iOS, Windows) — the free tier runs Whisper and Parakeet locally with no word caps and no account required. Trade-off: it's closed source, and its Pro tier (roughly $8-10/month) exists specifically to route audio through cloud models for higher accuracy — which reopens the exact subprocessor question this article started with. Only the free, local-only tier is local by architecture.
None of these are the "wrong" choice — they're real, working tools with genuine trade-offs on platform support, maturity, and what happens the moment you enable a cloud-accuracy option.
Inkvox: Local by Architecture, Built for Windows
Inkvox takes the same local-first approach, purpose-built for Windows. Whisper large-v3-turbo runs quantized directly on your own GPU through Vulkan — NVIDIA, AMD, or Intel — with a CPU fallback if you don't have a supported card. On a mid-range GPU like an RTX 3070, a sentence transcribes in roughly 0.3-0.4 seconds. Audio uploaded: 0 bytes. There's no account to create, no subprocessor list to check, and once the ~800 MB model finishes downloading, Inkvox runs fully offline, with support for 100+ languages.
Why that matters here: with Inkvox, there's no Privacy Mode toggle to go find, because there's no server in the data path to opt out of. The question this whole article raises — "where does my audio actually go?" — has a one-word answer: nowhere.
Inkvox is currently in free private beta for Windows 10/11, with the waitlist open. Inkvox Pro, a local LLM-powered rewrite layer, is in development, and macOS support is planned once the Windows beta hardens. If you handle regulated or privileged information for a living, see our dictation privacy guide for clinicians; if you're coming from the Mac side of this comparison, we cover a Windows-first alternative to SuperWhisper here. For the full breakdown of how Inkvox handles data end to end, see our privacy section.
Frequently Asked Questions
Is Wispr Flow safe to use?
Wispr Flow is a cloud dictation app: by its own documentation, transcription always happens on remote servers, routed through named subprocessors (Baseten, Soniox, OpenAI, Anthropic, Cerebras) and stored on AWS. That's a normal cloud SaaS architecture, and Wispr Flow discloses more about it than many peers. Whether it's "safe enough" depends on what you're dictating, and whether you're comfortable with audio leaving your device before it becomes text.
What does "zero data retention" mean in a dictation app?
For Wispr Flow specifically, it means Privacy Mode is turned on, which stops audio and transcripts from being stored or used for model training. It's off by default for standard accounts, and only locks on permanently if you sign the in-app HIPAA BAA. It's a data-handling policy applied after your audio reaches a server — not a guarantee that the audio never left your device.
Is there a local, offline alternative to Wispr Flow for Windows?
Yes. Inkvox runs Whisper large-v3-turbo directly on your GPU via Vulkan (NVIDIA, AMD, or Intel), with a CPU fallback, entirely on Windows 10/11 — audio never leaves the device, no account required, fully offline once the model downloads. It's currently in free private beta with an open waitlist. Cross-platform open-source options like OpenWhispr and Handy also run fully local on Windows.
Whatever tool you land on, the diligence is the same: ask where the audio goes, name the subprocessors, and check whether privacy is a default or a setting you have to find. If your answer is "nothing should leave my machine," join the Inkvox waitlist — Windows, 100% local, free beta.