← BlogFor developers

Free tier is uncapped, running locally

You used to write code. Now you write intent. And the tools you use to capture that intent still think you're trying to transcribe a meeting.

The bottleneck moved, but pricing didn't

A year ago, the bottleneck was typing speed. Now it's articulation clarity. When you work with Cursor, Claude Code, or GitHub Copilot, you're in a new loop: speak intent, review output, iterate. The faster you can externalize half-formed thoughts like architecture decisions, API designs, and error-handling edge cases, the faster the model builds and refines. Voice wins because you can think aloud. Your hands are free. Your rhythm is natural. You don't have to switch modes between thinking and typing.

But most voice-to-text tools charge you for the transcription itself. Wispr Flow costs $14 a month to unlock their free tier's word cap. Willow is $12. Superwhisper is $8.49. They all assume you're transcribing meeting notes or voice memos. That model works for occasional dictation. It breaks for what developers actually do: sit at 11pm with an architecture problem and voice-dump the full design doc.

The cap arrives mid-thought. You stop. You switch to typing. The flow shatters. The prose fragments. You clean it up the next morning, frustrated and exhausted.

Where it breaks, and why the IP risk matters

Marcus is a backend engineer at a fintech in Stockholm. Payment settlement systems, reconciliation pipelines, the kind of work where architecture decisions carry security weight. At 11pm, he's voice-dumping notes on a refactor: what the system looks like today, where the bottlenecks are, how to restructure without dropping transactions mid-flight.

Two minutes in, Wispr Flow's free tier cap hits. He stops. Switches to typing. Loses momentum. The thought scatters. Prose gets fragmented and dense because his hands can't keep up with his brain.

That's frustration. The second part is worse.

Marcus refuses cloud transcription entirely. He's uploading code architecture, internal naming conventions, security decisions, context around sensitive operations to a third party's servers. That's not caution. That's survival. The IP risk alone disqualifies any cloud-based tool, regardless of how good it is or how cheap it costs. Code thinking, security thinking, business model thinking: none of it leaves his device.

So Wispr, Willow, Superwhisper, all of them are off the table. They're all cloud. He's not paying for cloud transcription that locks him out of the tool he needs most.

Local Whisper changes the math

The reason every major voice tool charges $8 to $14 monthly is that cloud transcription has real costs. Audio processing, storage, API calls, network overhead, distribution, support. If you're running Whisper in the cloud, you're paying for compute and bandwidth on every single transcription. Of course it gets metered.

But Whisper Large-V3 is open-source. It hits 96.3% accuracy on LibriSpeech. Running it locally on modern hardware costs you zero per transcription. No API calls. No data leaving your device. No per-minute billing. The cost is literally zero.

Recitey runs Whisper on your device. No word cap. No metering. No uploads. You get uncapped dictation on the free tier because the structural cost to Recitey is zero. It's not a business model trick. It's not a loss leader. It's what the economics actually look like when you remove the cloud.

The trade-off is clarity. The Pro plan is where rewriting happens: after you've dictated your half-formed design doc or PR comment or Slack thread, Recitey can polish it into structured prose that reads like you spent an hour editing. That runs in the cloud and costs money. But the dictation itself, the part that keeps your thinking flowing, the part that doesn't need to be perfect yet, that's free and local.

You get the capture for free. You get the polish if you want it.

What actually changes in the workflow

Marcus voice-dumps the full design doc without hesitation now. Three minutes of architectural thinking, unbroken, into the system. Cursor's tab-complete does most of the cleanup work for him automatically, so even the rough draft is pretty readable. When it needs more polish, turning scattered notes into a Notion page the whole team can reference, he runs it through Pro and it's done in seconds.

No more fragmentation. No more stopping mid-thought to switch tools. No more code thinking leaving his device. The flow stays intact from capture through polish.

Most developers don't realize how much they lose when a tool interrupts the thinking loop. You're not just losing time. You're losing coherence. A half-formed thought that gets captured unbroken is much easier to refine than one that got paused, fragmented, and reassembled from memory the next morning.

The other shift is permission. Developers at companies with strong IP protection now have a voice tool they can use without running it by security first. It's local. It's uncapped. It doesn't need approval. They can use it right now.

Why this timing matters

The shift from typed code to voiced intent is real. It's already happening in the workflows of anyone using Claude, Copilot, or Cursor. But most voice tools were built for a different era: the era when transcription was the hard part, and everyone was expected to clean up the output before sharing it.

Tools built on that assumption price for transcription. They cap usage. They assume you're dictating occasionally and cleaning up the results carefully. That worked when people were transcribing recorded meetings.

It doesn't work when developers are using voice as a thinking tool. The new bottleneck isn't capture speed. It's keeping your thinking coherent and complete while three different models wait for your next prompt. Interrupting that flow, or forcing uploads of code thinking to the cloud, or metering usage at $14 a month, those aren't features of the old tools. They're constraints of the old business model.

Pricing based on cloud transcription cost no longer fits how developers actually work.

More posts
Keep reading

More like this.

  1. For developers

    Stop Hitting Word Caps in Your Design Docs

    Marcus is spending 11pm writing a design doc about payment settlement for a new feature. He's speaking clearly, the tool's...

  2. For developers

    The design doc that never got written

    It's 11 PM. Marcus, a backend engineer at a fintech in Stockholm, is voice-dictating a design doc for a new payment settlement...

  3. For developers

    Local transcription changes what you can dictate

    You wrote the design doc perfectly on voice, then scrolled up and realized 1400 words in, you'd stopped mid-sentence. Not...

All posts →