← BlogFor developers

The design doc that ran out of words

At 11pm, mid-explanation of your settlement system's state machine, the dictation caps out. Your cloud tool marks you at the limit. You switch back to typing, mid-thought, and the prose fractures. The next morning you're rewriting what should have been coherent.

For developers who've shifted to prompt-engineering and spec-writing as the real work, voice should be a way to move faster. But most voice tools were built for the previous workflow: capture quick memos, transcribe meetings. They meter words like it costs them money. It does, but that's their problem, not yours.

The workflow that broke cloud dictation

Marcus, a backend engineer at a fintech in Stockholm, spends his design doc time differently than it looked three years ago. He's not typing function signatures. He's dictating the intent behind them. He's walking through edge cases, architectural decisions, the shape of the data pipeline. That's what the LLM needs to understand before it writes the code.

His cloud dictation tool has a 2000-word-per-day cap on the free tier. Last quarter, he hit it on day eight of a design doc at 11pm. He switched back to typing. The voice-to-thought fluency broke. The next morning, the prose was fragmented, half voice, half typed, and he spent 30 minutes rewriting it into coherence.

Why cloud metering makes sense for the wrong use case

Wispr Flow charges $14/month and caps their free tier at 2000 words per day. Most cloud dictation tools follow a similar pattern. The variable cost is real, cloud processing isn't free, but the pricing reflects distribution and hosting, not the actual cost of Whisper running on your machine.

For developers, the use case is different. You're not capturing ten voice memos a day. You're drafting a long-form design doc in one sitting. You're explaining a bug in a Slack thread. You're describing what you want built in a Claude session. These are long-form, structured prose in professional contexts. The old pricing model breaks.

The technical difference that matters

Recitey runs Whisper locally. No upload, no cloud processing, no metering. The free tier has no word limit because the cost structure is structural: if the model runs on your device, there is no variable cost per transcription. The pro tier, the cloud rewrite that polishes rough voice output into production prose, that's where the cost sits. Not in capturing what you said, but in cleaning it up.

For Marcus, that's the difference. He refuses cloud transcription because his design docs reference internal code architecture and settlement flows. Once that voice leaves his laptop, it's data he doesn't control. With local Whisper, nothing leaves the device unless he chooses to send a final result to be rewritten. The IP stays local.

The flow you get back

Once the word-cap friction lifts, the workflow changes. Marcus now drafts a design doc by voice, uninterrupted, without watching a counter. When he gets stuck on clarity, he speaks the thought through three times until it makes sense, then lets the model clean it up. The dictation stays fast because there's no metering to think about. The thinking stays continuous because he's not context-switching back to the keyboard mid-sentence.

He still uses Cursor specifically because its tab-complete reduces how many voice rewrites he needs. Voice isn't a replacement for text editing. It's the thinking layer. Cursor handles the cleanup. Recitey handles the capture and the final polish. The tools do what they're good at.

What you trade

Local Whisper on a free tier is not cloud accuracy. It's accurate enough for intent. You'll get rough output. You'll need to edit. But you won't be re-recording because you hit a word limit in the middle of explaining a complex idea.

The trade is real: Recitey's free tier is not a finished product. The pro tier exists to make it one. But the free tier does what matters most: it captures without metering, it runs without leaving your device, and it works in whatever tool you're already using, Cursor, GitHub, Slack, Linear, wherever you write specifications and documentation.

For developers who've moved past typing code and into typing intent, that's the layer that mattered all along.

More posts
Keep reading

More like this.

  1. For developers

    Stop Hitting Word Caps in Your Design Docs

    Marcus is spending 11pm writing a design doc about payment settlement for a new feature. He's speaking clearly, the tool's...

  2. For developers

    The design doc that never got written

    It's 11 PM. Marcus, a backend engineer at a fintech in Stockholm, is voice-dictating a design doc for a new payment settlement...

  3. For developers

    Local transcription changes what you can dictate

    You wrote the design doc perfectly on voice, then scrolled up and realized 1400 words in, you'd stopped mid-sentence. Not...

All posts →